AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlog
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
Model Releases

Benchmarking Large Language Models for Safety Data Extraction

DGX agent

arXiv:2606.11204v1 Announce Type: new Abstract: Accurate extraction of structured information from Safety Data Sheets (SDS) remains challenging in industrial safety due to heterogeneous document forma

model-releasesarxiv-cs-cl
11 Jun 2026
Model Releases
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Beyond Compaction: Structured Context Eviction for Long-Horizon Agents

DGX agent

arXiv:2606.11213v1 Announce Type: new Abstract: We present Context Window Lifecycle (CWL), a context-management scheme that gives long-horizon LLM agents an effectively unbounded working horizon. As a

model-releasesarxiv-cs-cl
11 Jun 2026
Research

Beyond Fully Random Masking: Attention-Guided Denoising and Optimization for Diffusion Language Models

DGX agent

arXiv:2606.12273v1 Announce Type: new Abstract: Diffusion large language models (dLLMs) offer an efficient alternative to autoregressive models through parallel decoding, yet existing post-training me

researcharxiv-cs-cl
11 Jun 2026
Safety

Beyond Third-Person Audits: Situated Interaction Auditing for User-Centered LLM Bias Research

DGX agent

arXiv:2606.12247v1 Announce Type: cross Abstract: Research on bias in large language models (LLMs) has predominantly focused on third-person audits, which study how models represent or evaluate demogr

safetyarxiv-cs-cl
11 Jun 2026
Model Releases

BioMamba: Domain-Adaptive Biomedical Language Models

DGX agent

arXiv:2408.02600v3 Announce Type: replace Abstract: Background. Biomedical language models should improve performance on biomedical text while retaining general-language-modeling fluency. For Mamba-ba

model-releasesarxiv-cs-cl
11 Jun 2026
Agents

Breaking Entropy Bounds: Accelerating RL Training via MTP with Rejection Sampling

DGX agent

arXiv:2606.12370v1 Announce Type: cross Abstract: Reinforcement learning (RL) has become a key component in modern large language models, yet the rollout stage remains the key bottleneck in RL trainin

agentsarxiv-cs-cl
11 Jun 2026
Model Releases

Building Social World Models with Large Language Models

DGX agent

arXiv:2606.11482v1 Announce Type: cross Abstract: Understanding and predicting how social beliefs evolve in response to events -- from policy changes to scientific breakthroughs -- remains a fundament

model-releasesarxiv-cs-cl
11 Jun 2026
Safety

Can AI Reason Like an Urban Planner? Benchmarking Large Language Models Against Professional Judgment

DGX agent

arXiv:2606.11678v1 Announce Type: new Abstract: Problem, Research Strategy, and Findings: The rise of large language models (LLMs) raises a key question for urban planning: which forms of professional

safetyarxiv-cs-cl
11 Jun 2026
Research

Can News Predict the Market? Limits of Zero-Shot Financial NLP and the Role of Explainable AI

DGX agent

arXiv:2606.12210v1 Announce Type: new Abstract: Can financial news reliably predict short-term stock movements? Despite advances in large language models, this question remains unresolved. We revisit

researcharxiv-cs-cl
11 Jun 2026
Model Releases

Claw-SWE-Bench: A Benchmark for Evaluating OpenClaw-style Agent Harnesses on Coding Tasks

DGX agent

arXiv:2606.12344v1 Announce Type: cross Abstract: General-purpose agents such as OpenClaw are increasingly used as autonomous tool users, but their coding ability is difficult to measure under SWE-ben

model-releasesarxiv-cs-cl
11 Jun 2026
Safety

Compatibility-Aware Dynamic Fine-Tuning for Large Language Models

DGX agent

arXiv:2606.11206v1 Announce Type: new Abstract: Supervised Fine-Tuning (SFT) is the predominant paradigm for aligning large language models (LLMs), yet it suffers from optimization instability and lim

safetyarxiv-cs-cl
11 Jun 2026
Model Releases

Context-Aware Multimodal Claim Verification in Spoken Dialogues

DGX agent

arXiv:2606.11420v1 Announce Type: new Abstract: Every day, millions absorb claims from podcasts and streams that no fact-checker ever sees. Spoken misinformation is built through conversation, where c

model-releasesarxiv-cs-cl
11 Jun 2026
Research

Context-Driven Incremental Compression for Multi-Turn Dialogue Generation

DGX agent

arXiv:2606.12411v1 Announce Type: new Abstract: Modern conversational agents condition on an ever-growing dialogue history at each turn, incurring redundant attention and encoding costs that grow with

researcharxiv-cs-cl
11 Jun 2026
Model Releases

Debiasing Without Protected Attributes: Latent Concept Erasure from Textual Profiles

DGX agent

arXiv:2606.12088v1 Announce Type: new Abstract: Most fairness research in NLP assumes direct access to protected attributes such as gender, race, or nationality. In practice, however, such information

model-releasesarxiv-cs-cl
11 Jun 2026
Research

Decoding Multimodal Cues: Unveiling the Implicit Meaning Behind Hateful Videos

DGX agent

arXiv:2606.11953v1 Announce Type: new Abstract: Hateful videos have become prevalent on online platforms, highlighting an urgent need for effective detection. However, existing studies primarily focus

researcharxiv-cs-cl
11 Jun 2026
Applications

Detecting AI-Generated Content on Social Media with Multi-modal Language Models

DGX agent

arXiv:2606.11200v1 Announce Type: new Abstract: Generative AI has enabled the creation of photorealistic images and videos that are increasingly disseminated on social media, often used for spam, misi

applicationsarxiv-cs-cl
11 Jun 2026
Research

Detecting Sensitive Personal Information in Japanese Pre-Training Corpora for Large Language Models

DGX agent

arXiv:2606.12114v1 Announce Type: new Abstract: Sensitive personal information can appear in large-scale pre-training corpora for large language models (LLMs). Detecting and filtering such information

researcharxiv-cs-cl
11 Jun 2026
Research

Doc-to-Atom: Learning to Compile and Compose Memory Atoms

DGX agent

arXiv:2606.12400v1 Announce Type: new Abstract: Long input sequences are central to document understanding and multi-step reasoning in Large Language Models, yet the quadratic cost of attention makes

researcharxiv-cs-cl
11 Jun 2026
Safety

Dummy Backdoor as a Defense: Removing Unknown Backdoors via Shared Internal Mechanisms for Generative LLMs

DGX agent

arXiv:2606.11648v1 Announce Type: cross Abstract: Backdoor attacks pose a serious threat to the safety and reliability of Large Language Models (LLMs), as they cause models to behave normally on clean

safetyarxiv-cs-cl
11 Jun 2026
Model Releases

Energy-Efficient On-Device RAG on a Mobile NPU: System Design and Benchmark on Snapdragon X Elite

DGX agent

arXiv:2606.11257v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) pipelines are compute-intensive, combining embedding, retrieval, reranking, and large language model (LLM) generati

model-releasesarxiv-cs-cl
11 Jun 2026
Safety

Evaluating Bias in Phoneme-Based Automatic Speech Recognition Systems: An Analysis of IPA Transcription Models

DGX agent

arXiv:2606.11639v1 Announce Type: new Abstract: The popularization of automatic speech recognition (ASR) systems has increased exploration of the demographic biases related to race, age, gender, and a

safetyarxiv-cs-cl
11 Jun 2026
Model Releases

EverydayGPT: Confidence-Gated Routing for Efficient and Safe Hybrid GPT-RAG Conversational QA

DGX agent

arXiv:2606.11212v1 Announce Type: new Abstract: Standard Retrieval-Augmented Generation (RAG) pipelines route every query through retrieval and generation unconditionally, incurring unnecessary comput

model-releasesarxiv-cs-cl
11 Jun 2026
Safety

External Experience Serving in Production LLM Systems: A Deployment-Oriented Study of Quality-Cost Trade-offs

DGX agent

arXiv:2606.11806v1 Announce Type: new Abstract: Production LLM systems accumulate reusable operational experience, but the practical deployment issue is not merely whether such experience can help. It

safetyarxiv-cs-cl
11 Jun 2026
Research

Factions Within, Uncertain Across: Within-Document Reader Sub-Groups in Social Highlighting

DGX agent

arXiv:2606.11613v1 Announce Type: cross Abstract: When many people highlight the same document, is the crowd a single consensus, or is it internally structured into reader sub-groups that mark differe

researcharxiv-cs-cl
11 Jun 2026
Agents

Fanar-Sadiq: A Multi-Agent Architecture for Grounded Islamic QA

DGX agent

arXiv:2603.08501v3 Announce Type: replace Abstract: Large language models (LLMs) can answer religious knowledge queries fluently, yet they often hallucinate and misattribute sources, which is especial

agentsarxiv-cs-cl
11 Jun 2026
Research

Findings of the MAGMaR 2026 Shared Task

DGX agent

arXiv:2606.12295v1 Announce Type: cross Abstract: This overview paper presents the results of the shared task for the second workshop on Multimodal Augmented Generation via Multimodal Retrieval (MAGMa

researcharxiv-cs-cl
11 Jun 2026
Tutorials

FOCUS: DLLMs Know How to Tame Their Compute Bound

DGX agent

arXiv:2601.23278v2 Announce Type: replace-cross Abstract: Diffusion Large Language Models (DLLMs) offer a compelling alternative to Auto-Regressive models, but their deployment is constrained by high

tutorialsarxiv-cs-cl
11 Jun 2026
Research

FORT-Searcher: Synthesizing Shortcut-Resistant Search Tasks for Training Deep Search Agents

DGX agent

arXiv:2606.12087v1 Announce Type: new Abstract: Training deep search agents requires verifiable questions whose answers remain unavailable until sufficient evidence has been acquired through search. E

researcharxiv-cs-cl
11 Jun 2026
Model Releases

FronTalk: Benchmarking Front-End Development as Conversational Code Generation with Multi-Modal Feedback

DGX agent

arXiv:2601.04203v2 Announce Type: replace Abstract: We present FronTalk, a benchmark for front-end code generation that pioneers the study of a unique interaction dynamic: conversational code generati

model-releasesarxiv-cs-cl
11 Jun 2026
Model Releases

GraphInfer-Bench: Benchmarking LLM's Inference Capability on Graphs

DGX agent

arXiv:2606.11562v1 Announce Type: cross Abstract: Graph analysis underlies many applications whose answers cannot be looked up in a single record or retrieved along a path: laundering rings, drug repu

model-releasesarxiv-cs-cl
11 Jun 2026
Model Releases

GraspLLM: Towards Zero-Shot Generalization on Text-Attributed Graphs with LLMs

DGX agent

arXiv:2606.11898v1 Announce Type: new Abstract: Research on Text-Attributed Graphs (TAGs) has gained significant attention recently due to its broad applications across various real-world data scenari

model-releasesarxiv-cs-cl
11 Jun 2026
Research

Gumbel-BEARD: Automatic Layer Selection for Self-Supervised Adaptation of Whisper in Low-Resource Domains

DGX agent

arXiv:2606.11429v1 Announce Type: cross Abstract: Speech foundation models often struggle in low-resource domains due to domain mismatch and data scarcity. We propose Gumbel-BEARD, a domain adaptation

researcharxiv-cs-cl
11 Jun 2026
Model Releases

I Understand How You Feel: Enhancing Deeper Emotional Support Through Multilingual Emotional Validation in Dialogue System

DGX agent

arXiv:2606.11875v1 Announce Type: new Abstract: Emotional validation - explicitly acknowledging that a user's feelings make sense - has proven therapeutic value but has received little computational a

model-releasesarxiv-cs-cl
11 Jun 2026
Model Releases

Improving Cross-Format Robustness in Language Models with Multi-Format Training

DGX agent

arXiv:2606.11643v1 Announce Type: new Abstract: Large language models often remain sensitive to answer format: a question solved correctly in one form may fail in another semantically equivalent form.

model-releasesarxiv-cs-cl
11 Jun 2026
Research

Judging Against the Reference: Uncovering Knowledge-Driven Failures in LLM-Judges on QA Evaluation

DGX agent

arXiv:2601.07506v2 Announce Type: replace Abstract: While large language models (LLMs) are increasingly used as automatic judges for question answering (QA) and other reference-conditioned evaluation

researcharxiv-cs-cl
11 Jun 2026
Research

Kuramoto Attention: Synchronizing Self-Attention on the Torus

DGX agent

arXiv:2606.11585v1 Announce Type: cross Abstract: We introduce Kuramoto attention, a self-attention layer in which each hidden coordinate is an angle. The layer scores tokens by gated cosine similarit

researcharxiv-cs-cl
11 Jun 2026
Research

Language Shapes Mental Health Evaluations in Large Language Models

DGX agent

arXiv:2603.06910v2 Announce Type: replace Abstract: Multilingual large language models (LLMs) are increasingly used in socially sensitive mental health contexts, including support chatbots, screening,

researcharxiv-cs-cl
11 Jun 2026
Model Releases

LatticeBridge: Rare-Event Sequential Inference for Faithful Structured Sequence Synthesis

DGX agent

arXiv:2606.11203v1 Announce Type: new Abstract: Structured sequence generation often requires a model to satisfy several input-derived constraints in a single output. Standard decoding methods may ass

model-releasesarxiv-cs-cl
11 Jun 2026
Model Releases

LibriConvo: Simulating Conversations from Read Literature for ASR and Diarization

DGX agent

arXiv:2510.23320v2 Announce Type: replace-cross Abstract: We introduce LibriConvo, a synthetic conversational speech corpus for speaker diarization and automatic speech recognition (ASR), built by ins

model-releasesarxiv-cs-cl
11 Jun 2026
Model Releases

LifeSentence: Language models can encode human life course trajectories from longitudinal panel data

DGX agent

arXiv:2606.11220v1 Announce Type: new Abstract: Forecasting human life outcomes is important to gain insights into how individuals attain long and healthy lives. Conventional statistical approaches yi

model-releasesarxiv-cs-cl
11 Jun 2026
Research

Lius: Translation Model Based Instructional Lingustic Using Continual Instruction Tuning In Kupang Malay

DGX agent

arXiv:2606.11786v1 Announce Type: new Abstract: Large Language Models (LLMs) offer new potential for translation tasks but often experience performance degradation when handling low-resource languages

researcharxiv-cs-cl
11 Jun 2026
Model Releases

LLMpedia: A Transparent Framework to Materialize an LLM's Encyclopedic Knowledge at Scale

DGX agent

arXiv:2603.24080v2 Announce Type: replace Abstract: Benchmarks like MMLU suggest flagship language models approach factuality saturation above 90%. LLMpedia shows this picture is incomplete. We materi

model-releasesarxiv-cs-cl
11 Jun 2026
Applications

M4FC: a Multimodal, Multilingual, Multicultural, Multitask Real-World Fact-Checking Dataset

DGX agent

arXiv:2510.23508v3 Announce Type: replace Abstract: Existing real-world datasets for multimodal fact-checking have multiple limitations: they contain few instances, cover on only one or two languages,

applicationsarxiv-cs-cl
11 Jun 2026
Research

Massive Open-Vocabulary Keyword Spotting

DGX agent

arXiv:2606.11279v1 Announce Type: cross Abstract: Automatic speech recognition systems have been shown to under-perform when it comes to transcribing words rarely seen in the training data, namely spe

researcharxiv-cs-cl
11 Jun 2026
Agents

Measuring Epistemic Resilience of LLMs Under Misleading Medical Context

DGX agent

arXiv:2606.12291v1 Announce Type: new Abstract: Large language models (LLMs) now reach expert-level scores on medical licensing exams, encouraging the assumption that high scores imply safe medical ju

agentsarxiv-cs-cl
11 Jun 2026
Research

Measuring language complexity from hierarchical reuse of recurring patterns

DGX agent

arXiv:2606.11531v1 Announce Type: new Abstract: We introduce the ladderpath index as a measure of language complexity grounded in algorithmic information theory. It counts the minimum steps needed to

researcharxiv-cs-cl
11 Jun 2026
Safety

Measuring Semantic Progress in Multi-turn Dialogue via Information Gain

DGX agent

arXiv:2606.12332v1 Announce Type: new Abstract: Evaluating multi-turn dialogue is challenging because quality emerges across turns rather than within individual responses. We focus on a key dimension

safetyarxiv-cs-cl
11 Jun 2026
Model Releases

Multi-Agent Reasoning with Adaptive Worker Allocation for Stance Detection

DGX agent

arXiv:2606.11609v1 Announce Type: new Abstract: Stance detection requires identifying an author's position toward a target, often from short-form texts where stance is implicit, indirect, or rhetorica

model-releasesarxiv-cs-cl
11 Jun 2026
← Previous
1…4647484950…161
Next →