AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlog
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,646 results
Model Releases

TextNCA: Neural Cellular Automata for Language Modeling via Hierarchical Local Attention

DGX agent

arXiv:2608.02050v1 Announce Type: new Abstract: Can a strictly local, iterated, weight-shared computation primitive support language modelling, and which of those three properties actually drives the

model-releasesarxiv-cs-cl
4 Aug 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Research

The Holistic Storage of Verb+Up Phrases in Text-based and Audio-based Language Models

DGX agent

arXiv:2606.13993v2 Announce Type: replace Abstract: A crucial aspect of linguistic capability is the ability to trade off between stored representations and abstract knowledge: one must retrieve learn

researcharxiv-cs-cl
4 Aug 2026
Model Releases

The Learning Objective Governs Perceptual Narrowing: A Cross-Lingual, Layer-Wise, Ten-Seed Study of Self-Supervised Speech Encoders

DGX agent

arXiv:2608.00507v1 Announce Type: new Abstract: Perceptual narrowing---the developmental loss of non-native phoneme discrimination in the first year of life itep{werker1984}---is a canonical developme

model-releasesarxiv-cs-cl
4 Aug 2026
Research

The methodology of Constructing the Large-Scale Dataset for Detecting Presuicidal and Anti-Suicidal Signals in Social Media Texts in Russian

DGX agent

arXiv:2608.00497v1 Announce Type: new Abstract: The suicide is a terrifying act of a person who is misled by his own mental state. This problem arises across many countries. Sadly, Russia also has qui

researcharxiv-cs-cl
4 Aug 2026
Model Releases

The Role of Disfluencies in Speech Translation

DGX agent

arXiv:2608.02138v1 Announce Type: new Abstract: Current speech translation systems, including SpeechLLMs, are trained on cleaned text and tend to strip disfluencies like filled pauses and false starts

model-releasesarxiv-cs-cl
4 Aug 2026
Local Ai

Through the LENS: Local Geometric Decomposition of Vision-Language Model Representations

DGX agent

arXiv:2608.00561v1 Announce Type: cross Abstract: Vision-language models (VLMs) process image patches and text tokens in a shared residual stream, but the local geometry through which the two modaliti

local-aiarxiv-cs-cl
4 Aug 2026
Applications

TIDES: A Longitudinal Bilingual Dataset for Modeling Multi-Party Social Dynamics

DGX agent

arXiv:2608.01724v1 Announce Type: new Abstract: Group conversations are fundamental to human collaboration, yet standard large language models (LLMs) still struggle with the complexities of multi-part

applicationsarxiv-cs-cl
4 Aug 2026
Agents

Token-Native Storage: Read and Write in your Agent's Language

DGX agent

arXiv:2608.02376v1 Announce Type: cross Abstract: Search and database engines still store text as UTF-8, a format built for humans. But the systems that increasingly read and write that text (embedder

agentsarxiv-cs-cl
4 Aug 2026
Safety

Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning

DGX agent

arXiv:2608.01743v1 Announce Type: cross Abstract: Reinforcement learning (RL) has become a central paradigm for large language model (LLM) post-training, but optimization toward new objectives can deg

safetyarxiv-cs-cl
4 Aug 2026
Model Releases

Toward Robust LLM-Based Judges: Taxonomic Bias Evaluation and Debiasing Optimization

DGX agent

arXiv:2603.08091v2 Announce Type: replace Abstract: Large language model (LLM)-based judges are widely adopted for automated evaluation and reward modeling, yet their judgments are often affected by j

model-releasesarxiv-cs-cl
4 Aug 2026
Research

Towards a theory of morphology-driven marking in the lexicon: The case of the state

DGX agent

arXiv:2604.03422v2 Announce Type: replace Abstract: All languages have a noun category, but its realisation varies considerably. Depending on the language, semantic and/or morphosyntactic differences

researcharxiv-cs-cl
4 Aug 2026
Local Ai

TRACE-TS: Attribution-Grounded and Traceable Sensor-Language Reasoning for Human Activity Understanding

DGX agent

arXiv:2608.00200v1 Announce Type: cross Abstract: Wearable sensors capture fine-grained motion patterns that support rich behavioral understanding, yet most existing methods reduce these signals to ac

local-aiarxiv-cs-cl
4 Aug 2026
Research

Training-Free versus Training-Based Intent Classification in LLMs: Accuracy, Robustness, and Failure Modes

DGX agent

arXiv:2608.02415v1 Announce Type: new Abstract: Intent classification in Large Language Models (LLMs) involves categorizing user prompts into predefined classes. For instance, given a user prompt, the

researcharxiv-cs-cl
4 Aug 2026
Research

TRAM: Enhancing Multimodal Reasoning with Trajectory-Derived Auxiliary Memory

DGX agent

arXiv:2608.01922v1 Announce Type: new Abstract: Multimodal Large Reasoning Models (MLRMs) have achieved strong performance on tasks requiring visual understanding and multi-step inference. However, as

researcharxiv-cs-cl
4 Aug 2026
Research

Transformers perform adaptive partial pooling

DGX agent

arXiv:2602.03980v2 Announce Type: replace Abstract: Any language model must decide what to say in novel contexts based on information from similar contexts. But what about contexts that are not novel

researcharxiv-cs-cl
4 Aug 2026
Model Releases

TreeProbe : A Tibetan Medicine Benchmark for Cultural Bias in LLMs

DGX agent

arXiv:2608.00640v1 Announce Type: new Abstract: Large language models are increasingly viewed as a potential means of mitigating global health inequities, yet their outputs often reflect dominant high

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

TrimMoE A communication aware and adaptive depth framework for distributed edge inference

DGX agent

arXiv:2608.00573v1 Announce Type: cross Abstract: Serving Mixture-of-Experts (MoE) large language models across distributed edge servers is bottlenecked by the cross-server expert transmission. The ex

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

Trustworthiness Costs of Domain Adaptation in Small Language Models:A Cross-Architecture Empirical Study

DGX agent

arXiv:2608.00042v1 Announce Type: new Abstract: Domain adaptation of small language models (SLMs) has emerged as a practical strategy for deploying capable NLP systems in resource-constrained, high-st

model-releasesarxiv-cs-cl
4 Aug 2026
Safety

TS-Reasoner: Aligning Time Series Foundation Models with LLM Reasoning

DGX agent

arXiv:2510.03519v2 Announce Type: replace Abstract: Time series reasoning is crucial to decision-making in diverse domains, including finance, energy, and scientific discovery. While existing time ser

safetyarxiv-cs-cl
4 Aug 2026
Model Releases

Two-Stage Bengali Sentiment Classification: Domain Adaptation Through Continual Learning and Parameter-Efficient Fine-Tuning

DGX agent

arXiv:2608.01471v1 Announce Type: new Abstract: Understanding sentiment in low-resource languages remains a key challenge for Natural Language Processing (NLP), particularly when domain-specific data

model-releasesarxiv-cs-cl
4 Aug 2026
Agents

UEmbed: Unified Sparse and Dense Multimodal Embeddings

DGX agent

arXiv:2608.02583v1 Announce Type: cross Abstract: Sparse retrieval underpins modern search systems, from web search to retrieval-augmented generation. Existing work has introduced Learned Sparse Retri

agentsarxiv-cs-cl
4 Aug 2026
Research

Understanding Sparse Attention Selectivity in Long-Context Foundation Models via Counterfactual Evaluation

DGX agent

arXiv:2608.01676v1 Announce Type: new Abstract: Sparse attention is widely deployed in long-context serving stacks, yet no framework audits how discarding blocks changes the influence of specific cont

researcharxiv-cs-cl
4 Aug 2026
Research

UniHEAR: Unified Heterogeneous-Source Attentive Retrieval for Knowledge-Based Visual Question Answering

DGX agent

arXiv:2608.01147v1 Announce Type: cross Abstract: Knowledge-Based Visual Question Answering (KB-VQA) requires retrieving relevant entity knowledge from external sources to answer visually grounded que

researcharxiv-cs-cl
4 Aug 2026
Safety

Unleashing the Potential of Large Language Models: A Blueprint for Real-Time, Enterprise-Ready Deployments

DGX agent

arXiv:2608.00419v1 Announce Type: cross Abstract: Large language models deployed in real-time, regulated settings face knowledge staleness, catastrophic forgetting, hallucination, and weak feedback lo

safetyarxiv-cs-cl
4 Aug 2026
Research

Unpacking Hateful Memes: Presupposed Context and False Claims

DGX agent

arXiv:2510.09935v2 Announce Type: replace Abstract: While memes are often humorous, they are frequently used to disseminate hate, causing serious harm to individuals and society. Current approaches to

researcharxiv-cs-cl
4 Aug 2026
Research

Unsupervised Multidomain Approaches to Named Entity Recognition with Small Datasets

DGX agent

arXiv:2608.00984v1 Announce Type: new Abstract: This paper explores the challenges and the methodologies associated with learning quality representations in scenarios with unlabelled small or limited

researcharxiv-cs-cl
4 Aug 2026
Agents

V-Mem: Modality-Routed Retrieval for Long-Term Multimodal Agentic Memory

DGX agent

arXiv:2608.01543v1 Announce Type: cross Abstract: Interaction between users and LLM agents is increasingly multimodal: conversations interleave text with images, and a later question may target either

agentsarxiv-cs-cl
4 Aug 2026
Research

Verification Without Sufficiency: Per-Chunk Filtering Fails on Multi-Hop RAG, and Decomposition Repairs It

DGX agent

arXiv:2608.00585v1 Announce Type: new Abstract: Verification for retrieval-augmented generation usually scores each retrieved chunk and drops the ones that fail. We show this cannot work for multi-hop

researcharxiv-cs-cl
4 Aug 2026
Safety

Verifier-Induced Support Reshaping in On-Policy Optimization

DGX agent

arXiv:2608.00220v1 Announce Type: cross Abstract: We show that on-policy reinforcement learning with verifiable rewards (RLVR) can improve the current objective while making successful behaviors for l

safetyarxiv-cs-cl
4 Aug 2026
Research

Visualising Information Flow in Word Embeddings with Diffusion Tensor Imaging

DGX agent

arXiv:2601.05713v2 Announce Type: replace Abstract: Understanding how large language models (LLMs) represent natural language is a central challenge in natural language processing (NLP) research. Many

researcharxiv-cs-cl
4 Aug 2026
Model Releases

What Makes Position Zero Special? A Mechanistic Study of Position Zero Attention Sinks in LLMs

DGX agent

arXiv:2603.06591v2 Announce Type: replace-cross Abstract: Transformers frequently allocate disproportionate attention to specific tokens, a phenomenon known as attention sinks. Causal large language m

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

What Transfers from Text to Vision? Capability Scaling Laws and Transfer Dynamics for VLMs

DGX agent

arXiv:2608.00013v1 Announce Type: new Abstract: Choosing the right large language model (LLM) backbone is the most consequential decision when building a vision-language model (VLM), yet it remains fu

model-releasesarxiv-cs-cl
4 Aug 2026
Research

When LLM Essays Outscore Student Essays: What a Korean Writing Rubric Rewards and Where Readers Disagree

DGX agent

arXiv:2601.19913v4 Announce Type: replace Abstract: LLMs now help students plan, draft, and revise essays. Educational assessment therefore faces a basic question: how should student and LLM writing b

researcharxiv-cs-cl
4 Aug 2026
Agents

When Only the Final Text Survives: Implicit Execution Tracing for Multi-Agent Auditing

DGX agent

arXiv:2603.17445v5 Announce Type: replace-cross Abstract: When a multi-agent system produces an incorrect or harmful answer, who is accountable if execution logs and agent identifiers are unavailable?

agentsarxiv-cs-cl
4 Aug 2026
Model Releases

When Retrieval Helps and Distracts: Evaluating Evidence-Generating LLMs for Biomedical Claim Verification

DGX agent

arXiv:2608.01409v1 Announce Type: new Abstract: Biomedical fact-checking systems must do more than predict whether a claim is supported, contradicted, or unaddressed: they should also produce evidence

model-releasesarxiv-cs-cl
4 Aug 2026
Research

When Words Divide: Diachronic Ideological Polarization in Political Discourse on Social Media

DGX agent

arXiv:2608.01176v1 Announce Type: new Abstract: Political polarization has become a defining feature of online discourse, yet its long-term evolution remains poorly understood. We present a longitudin

researcharxiv-cs-cl
4 Aug 2026
Research

Where did the ambiguity go? Examining how multimodal models interpret polysemous words

DGX agent

arXiv:2608.00410v1 Announce Type: cross Abstract: Human language is highly polysemous. Many common words (e.g., 'bank' or 'palm') carry several distinct meanings that shape what humans communicate and

researcharxiv-cs-cl
4 Aug 2026
Safety

Who Should Be Generated? Justifying Demographic Targets in Open-Ended Generation

DGX agent

arXiv:2608.02551v1 Announce Type: cross Abstract: Fairness evaluation concerns not only what a model produces, but also what its outputs ought to be compared against. When a model generates 'a CEO in

safetyarxiv-cs-cl
4 Aug 2026
Research

Why LLMs Give In: Conversational Factors and Reasoning Behind Medical Sycophancy

DGX agent

arXiv:2608.01017v1 Announce Type: new Abstract: A language model that abandons a correct medical answer under user pushback is more dangerous than one that was simply wrong, because it lends the credi

researcharxiv-cs-cl
4 Aug 2026
Model Releases

Writing-System-Level Tokenizer Adaptation for Byte-Level BPE

DGX agent

arXiv:2608.00582v1 Announce Type: new Abstract: Pretrained byte-level BPE tokenizers can segment underrepresented languages inefficiently. Replacing a tokenizer changes the meaning of nearly every tok

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

XL-DocBench: Benchmarking Evidence-Grounded Extra-Long Document Understanding

DGX agent

arXiv:2608.00036v1 Announce Type: new Abstract: Real-world document tasks often ask professionals to answer questions from annual reports, regulations, clinical guidelines, and technical manuals that

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

Are the Financial Reasoning from LLMs Credible? A Real World Test over Long-Horizon Statements

DGX agent

arXiv:2607.28661v1 Announce Type: new Abstract: Do Large Language Models (LLMs) possess genuine structural reasoning, or merely rely on surface-level pattern matching? The financial domain, demanding

model-releasesarxiv-cs-cl
3 Aug 2026
Research

Authorship Verification of Transcribed German-Language Videos

DGX agent

arXiv:2607.29168v1 Announce Type: new Abstract: Authorship Verification (AV) represents an important subfield of digital text forensics and addresses the fundamental question of whether two texts were

researcharxiv-cs-cl
3 Aug 2026
Model Releases

Benchmarks Are Not Validation: A System-Level View of Financial LLM Applications

DGX agent

arXiv:2607.28840v1 Announce Type: new Abstract: Large language models are increasingly deployed in financial applications that combine retrieval, proprietary data, tool use, orchestration logic, monit

model-releasesarxiv-cs-cl
3 Aug 2026
Research

BLADE: Boundary-Expanded and Layer-Adaptive Dynamic Exit for Efficient LLM Reasoning

DGX agent

arXiv:2607.28966v1 Announce Type: new Abstract: Large language models often improve task performance by generating long reasoning traces, but the resulting computation is frequently wasted on redundan

researcharxiv-cs-cl
3 Aug 2026
Safety

Bridging the Question-Answer Gap in Retrieval-Augmented Generation: Hypothetical Prompt Embeddings

DGX agent

arXiv:2607.29402v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) systems synergize retrieval mechanisms with generative language models to enhance the accuracy and relevance of r

safetyarxiv-cs-cl
3 Aug 2026
Model Releases

Can LLMs Really Understand Item Difficulty Levels? Implications for Automated Item Generation Using LLMs

DGX agent

arXiv:2607.28634v1 Announce Type: new Abstract: The estimation of item difficulty plays a key role in both formative assessment and large-scale high-stakes summative assessments. This study explores h

model-releasesarxiv-cs-cl
3 Aug 2026
Safety

Can Zero-Shot LLMs Predict Child Malnutrition? A Fairness and Temporal Robustness Study

DGX agent

arXiv:2607.29082v1 Announce Type: new Abstract: Child malnutrition remains a major public health challenge in low- and middle-income countries, particularly in South Asia, where early identification o

safetyarxiv-cs-cl
3 Aug 2026
← Previous
1…1314151617…160
Next →