AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,759 results
Model Releases

CARTE: A Benchmark for Mapping Language Model Knowledge Across France

DGX agent

arXiv:2606.01995v1 Announce Type: new Abstract: We introduce CARTE 1 (Culturally Anchored Regional-Territorial Evaluation), a multiplechoice benchmark for evaluating the ability of large language mode

model-releasesarxiv-cs-cl
2 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

Challenger at MultiPRIDE: Is It Hate Speech or Reclaimed?

DGX agent

arXiv:2606.01298v1 Announce Type: new Abstract: The spread of hate speech has become increasingly harmful in modern digital environments, particularly on social networking platforms. While recent adva

researcharxiv-cs-cl
2 Jun 2026
Research

Characterizing the Effect of Noise in Language Generation in the Limit

DGX agent

arXiv:2601.21237v2 Announce Type: replace-cross Abstract: Kleinberg and Mullainathan recently proposed a formal framework for studying the phenomenon of language generation, called language generation

researcharxiv-cs-cl
2 Jun 2026
Model Releases

Child-directed speech facilitates production, not comprehension, in BabyLMs

DGX agent

arXiv:2606.01045v1 Announce Type: new Abstract: Recent studies suggest that child-directed speech is not conducive to language learning in BabyLMs. However, current evaluations focus predominantly on

model-releasesarxiv-cs-cl
2 Jun 2026
Research

Chunking Methods on Retrieval-Augmented Generation - Effectiveness Evaluation Against Computational Cost and Limitations

DGX agent

arXiv:2606.00881v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) has demonstrated significant capabilities in enhancing the performance of Large Language Models (LLMs). One of the

researcharxiv-cs-cl
2 Jun 2026
Model Releases

Citation Grounding: Detecting and Reducing LLM Citation Hallucinations via Legal Citation Graphs

DGX agent

arXiv:2606.00898v1 Announce Type: new Abstract: Large language models systematically hallucinate legal citations -- fabricating statute references, citing repealed provisions, and confusing jurisdicti

model-releasesarxiv-cs-cl
2 Jun 2026
Safety

ClinTutor-R1: Advancing Scalable and Robust One-to-Many Alignment in Clinical Socratic Education

DGX agent

arXiv:2512.05671v2 Announce Type: replace Abstract: While Large Language Models (LLMs) have achieved remarkable success in dyadic (one-on-one) instruction, they face significant challenges in One-to-M

safetyarxiv-cs-cl
2 Jun 2026
Agents

Code2Math: Can Your Code Agent Effectively Evolve Math Problems Through Exploration?

DGX agent

arXiv:2603.03202v3 Announce Type: replace Abstract: As large language models (LLMs) advance their mathematical capabilities toward the IMO and research level, the scarcity of challenging, high-quality

agentsarxiv-cs-cl
2 Jun 2026
Research

Cognitive-Linguistic Indicators of Depression in Online Communities: Analysed by DistilBERT and Holographic Reduced Representation

DGX agent

arXiv:2606.00026v1 Announce Type: new Abstract: This paper investigates whether combining cognitively grounded linguistic features with transformer-based embeddings improves automated detection of dep

researcharxiv-cs-cl
2 Jun 2026
Research

Confidence-Adaptive SwiGLU for Mixture-of-Experts

DGX agent

arXiv:2606.00761v1 Announce Type: cross Abstract: SwiGLU has become a standard gated activation in modern Transformer MLPs, yet its gate sharpness -- the smoothness and selectivity of the gating funct

researcharxiv-cs-cl
2 Jun 2026
Model Releases

ContinuousBench: Can Differentially Private Synthetic Text Improve Capabilities?

DGX agent

arXiv:2606.01849v1 Announce Type: cross Abstract: Differentially private (DP) text synthesis promises to unlock sensitive corpora for model training, but it remains unclear whether DP synthetic data t

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Correcting Gradient-Based Circuit Localization via Interaction-Aware Backpropagation

DGX agent

arXiv:2505.17630v4 Announce Type: replace Abstract: Circuit localization methods aim to identify the subset of model components responsible for specific behaviors in large language models, enabling de

model-releasesarxiv-cs-cl
2 Jun 2026
Research

Cost-Aware Diffusion Draft Trees for Speculative Decoding

DGX agent

arXiv:2606.01813v1 Announce Type: new Abstract: Speculative decoding accelerates inference by having a lightweight drafter propose tokens verified in parallel by the target language model. Block diffu

researcharxiv-cs-cl
2 Jun 2026
Model Releases

CRAB-Bench: Evaluating LLM Agents under Complex Task Dependencies and Human-aligned User Simulation

DGX agent

arXiv:2606.01815v1 Announce Type: new Abstract: Evaluating LLM agents in realistic service scenarios requires complex task dependencies, imperfect user behavior, and an evaluation that accommodates mu

model-releasesarxiv-cs-cl
2 Jun 2026
Tutorials

CRAFTQA: A Code-Driven Adaptive Framework for Complex Structured Data Reasoning

DGX agent

arXiv:2606.02170v1 Announce Type: new Abstract: Real-world scenarios involve massive heterogeneous structured data (e.g., tables, knowledge graphs), making effective reasoning over such diverse data i

tutorialsarxiv-cs-cl
2 Jun 2026
Model Releases

CRAM: Centroid-Routing and Adaptive MoE for Multimodal Continual Instruction Tuning

DGX agent

arXiv:2606.02502v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) unify heterogeneous vision-language tasks under a shared generative framework via instruction tuning, yet real-

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Cross-Environment Neural Reranking for Sample-Efficient Action Selection in Text-Based Agents

DGX agent

arXiv:2606.02204v1 Announce Type: new Abstract: Large language model agents achieve strong performance on text-based benchmarks but incur prohibitive inference costs, motivating the use of compact neu

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Cross-Generational Transfer of Adversarial Attacks Reveals Non-Monotonic Safety Alignment in LLMs

DGX agent

arXiv:2606.00813v1 Announce Type: cross Abstract: Safety alignment in LLMs does not improve monotonically across model generations. Studying four generations of Google's Gemma family (7B-31B) with qua

model-releasesarxiv-cs-cl
2 Jun 2026
Research

Cross-lingual Self-Consistency for Multilingual Reasoning with Language Models

DGX agent

arXiv:2606.01464v1 Announce Type: new Abstract: Despite expanding their multilingual coverage, the advanced reasoning capabilities of LLMs remain largely confined to a few high-resource languages like

researcharxiv-cs-cl
2 Jun 2026
Model Releases

CultureForest: Understanding and Evaluating Cultural Norm Grounded Reasoning in LLMs

DGX agent

arXiv:2606.01879v1 Announce Type: new Abstract: Existing research largely reduces cultural intelligence in LLMs to a knowledge-level problem, overlooking whether models can effectively utilize their a

model-releasesarxiv-cs-cl
2 Jun 2026
Research

CURP: Codebook-based Continuous User Representation for Personalized Generation with LLMs

DGX agent

arXiv:2602.00742v2 Announce Type: replace Abstract: User modeling characterizes individuals through their preferences and behavioral patterns to enable personalized simulation and generation with Larg

researcharxiv-cs-cl
2 Jun 2026
Model Releases

DECK: A Consistency x Confidence Taxonomy of LLM Hallucinations

DGX agent

arXiv:2606.02289v1 Announce Type: new Abstract: Existing hallucination taxonomies classify LLM errors by what is wrong with the output -- memorised misconceptions, reasoning failures, fluent fabricati

model-releasesarxiv-cs-cl
2 Jun 2026
Research

Decoding in Order-Agnostic Language Models: Chain-Rule Deviation and Uniform Spreading

DGX agent

arXiv:2606.00997v1 Announce Type: new Abstract: Order-agnostic language models (OALMs), including discrete diffusion language models (dLLMs), are trained to predict masked tokens under arbitrary condi

researcharxiv-cs-cl
2 Jun 2026
Safety

Decomposed On-Policy Distillation for Vision-Language Reasoning: Steering Gradients for Visual Grounding

DGX agent

arXiv:2606.00564v1 Announce Type: cross Abstract: While on-policy distillation offers dense supervision for training small reasoning models, its optimization dynamics in the multimodal domain remain u

safetyarxiv-cs-cl
2 Jun 2026
Tutorials

Deep networks learn to parse uniform-depth context-free languages from local statistics

DGX agent

arXiv:2602.06065v3 Announce Type: replace-cross Abstract: Understanding how the structure of language can be learned from sentences alone is a central question in both cognitive science and machine le

tutorialsarxiv-cs-cl
2 Jun 2026
Model Releases

Deep Research as Rubric for Reinforcement Learning

DGX agent

arXiv:2606.01091v1 Announce Type: new Abstract: Open-ended reasoning and long-form generation tasks lack reliable automatic verification signals for reward-based policy optimization. Rubrics offer a p

model-releasesarxiv-cs-cl
2 Jun 2026
Research

DeSQ: Decomposition-based SPARQL Query Generation

DGX agent

arXiv:2606.00203v1 Announce Type: new Abstract: Dominant approaches to Knowledge Base Question Answering (KBQA) fall into two categories. First is the generation of a formal query that suffers from br

researcharxiv-cs-cl
2 Jun 2026
Research

DFlare: Scaling Up Draft Capacity for Block Diffusion Speculative Decoding

DGX agent

arXiv:2606.02091v1 Announce Type: new Abstract: Block diffusion speculative decoding accelerates LLM inference by predicting all tokens within a block simultaneously for the target model to verify in

researcharxiv-cs-cl
2 Jun 2026
Research

Digging Up Citations: FOSSIL, a Dataset and Workflow for Reference Extraction in Law and the Humanities

DGX agent

arXiv:2606.01109v1 Announce Type: cross Abstract: Citation extraction tools are designed for the structured end-of-document bibliographies of the natural sciences, but law and humanities scholarship c

researcharxiv-cs-cl
2 Jun 2026
Model Releases

Disentangling Similarity and Relatedness in Topic Models

DGX agent

arXiv:2603.10619v2 Announce Type: replace Abstract: The recent success of large pre-trained language models (PLMs) has motivated their integration into topic modeling. However, PLM-augmented topic mod

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Do Gender Cues Affect LLM Value Trade-offs? Evidence from a Controlled Decision Benchmark

DGX agent

arXiv:2606.02214v1 Announce Type: new Abstract: Large language models are increasingly used in value-sensitive decision settings, where irrelevant demographic cues should not alter judgments. We const

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Do Text Edits Generalize to Visual Generation? Benchmarking Cross-Modal Knowledge Editing in UMMs

DGX agent

arXiv:2606.00477v1 Announce Type: new Abstract: Unified multimodal models (UMMs) have emerged as a promising paradigm for general-purpose multimodal intelligence. As they are deployed in real-world ap

model-releasesarxiv-cs-cl
2 Jun 2026
Local Ai

Don't Read Everything: A Curvature-Conditioned Query for Linear Attention

DGX agent

arXiv:2606.01294v1 Announce Type: new Abstract: Linear attention reduces the quadratic cost of softmax attention by maintaining a recurrent fast-weight state, but it consistently lags on in-context re

local-aiarxiv-cs-cl
2 Jun 2026
Model Releases

DrugClaw and DrugAudit: A Primary-Source-Grounded Agent and Authority-Aware Benchmark for Drug-Information Question Answering

DGX agent

arXiv:2606.01434v1 Announce Type: new Abstract: Drug-information question answering is a high-stakes setting where hallucinated facts can mislead clinical decision-making and the provenance of each ci

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Efficient RAG with Intent-Aware Retrieval and Semantics-Preserving Chunking

DGX agent

arXiv:2606.01240v1 Announce Type: new Abstract: The demand for powerful instruction following and reasoning capability of large language models (LLMs) has promoted rapid development of retrieval-augme

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Empathy Applicability Modeling for General Health Queries

DGX agent

arXiv:2601.09696v2 Announce Type: replace Abstract: LLMs are increasingly being integrated into clinical workflows, yet they often lack clinical empathy, an essential aspect of effective doctor-patien

model-releasesarxiv-cs-cl
2 Jun 2026
Research

Encoded but Not Routed: Explaining the Table-Chart Gap in Scientific Claim Verification

DGX agent

arXiv:2606.01679v1 Announce Type: new Abstract: Multimodal LLMs are increasingly used to assist scientific peer review, where a core requirement is verifying whether claims in a paper are supported by

researcharxiv-cs-cl
2 Jun 2026
Model Releases

Escaping the Mode Lottery: Multi-Response Training Improves Language Model Generalization

DGX agent

arXiv:2606.00544v1 Announce Type: cross Abstract: Modern language-model fine-tuning typically pairs each prompt with a single response, even though many prompts admit multiple valid completions. This

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Evaluating the Reversal Curse in Model Editing

DGX agent

arXiv:2310.10322v3 Announce Type: replace Abstract: Large language models (LLMs) are prone to hallucinate unintended text due to false or outdated knowledge. Since retraining LLMs is resource intensiv

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

ExpWeaver: LLM Agents Learn from Experience via Latent RAG

DGX agent

arXiv:2606.01041v1 Announce Type: new Abstract: Experience learning has achieved promising results in enhancing LLM agent planning and reasoning by integrating past interactions as reusable knowledge.

model-releasesarxiv-cs-cl
2 Jun 2026
Hardware

Eyettention II: A Dual-Sequence Architecture for Modeling Fixation Location, Within-Word Landing Position, and Fixation Duration in Reading

DGX agent

arXiv:2606.01964v1 Announce Type: new Abstract: The way our eyes move while reading provides valuable insights into both the reader's cognitive processes and the properties of the text. In particular,

hardwarearxiv-cs-cl
2 Jun 2026
Model Releases

FigSIM: A Dataset for Fine-grained Suicide Severity and Figurative Language in Suicide Memes

DGX agent

arXiv:2606.02523v1 Announce Type: new Abstract: Suicide memes are memes used to express suicide-related thoughts or comment on suicide-related issues. Suicide memes are increasingly common on social m

model-releasesarxiv-cs-cl
2 Jun 2026
Tutorials

Finding What Matters: Anchoring Context Knowledge with Evolving Indices for Iterative Retrieval

DGX agent

arXiv:2601.16462v2 Announce Type: replace Abstract: Retrieval-Augmented Generation (RAG) has become a dominant paradigm for mitigating hallucinations in Large Language Models (LLMs) by incorporating e

tutorialsarxiv-cs-cl
2 Jun 2026
Applications

Fine-Tuning Without Forgetting In-Context Learning: A Theoretical Analysis of Linear Attention Models

DGX agent

arXiv:2602.23197v2 Announce Type: replace Abstract: Transformer-based large language models exhibit in-context learning, enabling adaptation to downstream tasks via few-shot prompting with demonstrati

applicationsarxiv-cs-cl
2 Jun 2026
Model Releases

FineVerify: Scaling Test-Time Compute with Fine-Grained Self-Verification for Agentic Search

DGX agent

arXiv:2606.00660v1 Announce Type: new Abstract: Agentic search requires language model agents to explore many sources and answer complex information-seeking questions. Scaling test-time compute is a p

model-releasesarxiv-cs-cl
2 Jun 2026
Research

ForesightKV: Optimizing KV Cache Eviction for Reasoning Models by Learning Long-Term Contribution

DGX agent

arXiv:2602.03203v2 Announce Type: replace Abstract: Recently, large language models (LLMs) have shown remarkable reasoning abilities by producing long reasoning traces. However, as the sequence length

researcharxiv-cs-cl
2 Jun 2026
Research

French parsing enhanced with a word clustering method based on a syntactic lexicon

DGX agent

arXiv:2606.00634v1 Announce Type: new Abstract: This article evaluates the integration of data extracted from a French syntactic lexicon, the Lexicon-Grammar (Gross, 1994), into a probabilistic parser

researcharxiv-cs-cl
2 Jun 2026
Research

From Empathy to Personalized Empathy: Adapting Empathetic Strategies to Individual Users

DGX agent

arXiv:2606.00728v1 Announce Type: new Abstract: As Large Language Models (LLMs) are increasingly deployed in long-term interactions with users, empathy has become an increasingly important capability.

researcharxiv-cs-cl
2 Jun 2026
← Previous
1…6162636465…162
Next →