AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlog
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,646 results
Research

CoG: Controllable Graph Reasoning via Relational Blueprints and Failure-Aware Refinement over Knowledge Graphs

DGX agent

arXiv:2601.11047v2 Announce Type: replace Abstract: Large Language Models (LLMs) have demonstrated remarkable reasoning capabilities but often grapple with reliability challenges like hallucinations.

researcharxiv-cs-cl
15 Apr 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Safety

Compiling Activation Steering into Weights via Null-Space Constraints for Stealthy Backdoors

DGX agent

arXiv:2604.12359v1 Announce Type: cross Abstract: Safety-aligned large language models (LLMs) are increasingly deployed in real-world pipelines, yet this deployment also enlarges the supply-chain atta

safetyarxiv-cs-cl
15 Apr 2026
Model Releases

CompliBench: Benchmarking LLM Judges for Compliance Violation Detection in Dialogue Systems

DGX agent

arXiv:2604.12312v1 Announce Type: new Abstract: As Large Language Models (LLMs) are increasingly deployed as task-oriented agents in enterprise environments, ensuring their strict adherence to complex

model-releasesarxiv-cs-cl
15 Apr 2026
Safety

ContextLens: Modeling Imperfect Privacy and Safety Context for Legal Compliance

DGX agent

arXiv:2604.12308v1 Announce Type: new Abstract: Individuals' concerns about data privacy and AI safety are highly contextualized and extend beyond sensitive patterns. Addressing these issues requires

safetyarxiv-cs-cl
15 Apr 2026
Research

CoRoVA: Compressed Representations for Vector-Augmented Code Completion

DGX agent

arXiv:2510.19644v2 Announce Type: replace Abstract: Retrieval-augmented generation has emerged as one of the most effective approaches for code completion enhancement, especially when repository-level

researcharxiv-cs-cl
15 Apr 2026
Research

Do Transformers Use their Depth Adaptively? Evidence from a Relational Reasoning Task

DGX agent

arXiv:2604.12426v1 Announce Type: cross Abstract: We investigate whether transformers use their depth adaptively across tasks of increasing difficulty. Using a controlled multi-hop relational reasonin

researcharxiv-cs-cl
15 Apr 2026
Model Releases

Do VLMs Truly 'Read' Candlesticks? A Multi-Scale Benchmark for Visual Stock Price Forecasting

DGX agent

arXiv:2604.12659v1 Announce Type: cross Abstract: Vision-language models(VLMs) are increasingly applied to visual stock price forecasting, yet existing benchmarks inadequately evaluate their understan

model-releasesarxiv-cs-cl
15 Apr 2026
Research

E2LLM: Encoder Elongated Large Language Models for Long-Context Understanding and Reasoning

DGX agent

arXiv:2409.06679v3 Announce Type: replace Abstract: Processing long contexts is increasingly important for Large Language Models (LLMs) in tasks like multi-turn dialogues, code generation, and documen

researcharxiv-cs-cl
15 Apr 2026
Research

Efficient Inference for Large Vision-Language Models: Bottlenecks, Techniques, and Prospects

DGX agent

arXiv:2604.05546v2 Announce Type: replace Abstract: Large Vision-Language Models (LVLMs) enable sophisticated reasoning over images and videos, yet their inference is hindered by a systemic efficiency

researcharxiv-cs-cl
15 Apr 2026
Model Releases

Empirical Evaluation of PDF Parsing and Chunking for Financial Question Answering with RAG

DGX agent

arXiv:2604.12047v1 Announce Type: new Abstract: PDF files are primarily intended for human reading rather than automated processing. In addition, the heterogeneous content of PDFs, such as text, table

model-releasesarxiv-cs-cl
15 Apr 2026
Research

Enhance-then-Balance Modality Collaboration for Robust Multimodal Sentiment Analysis

DGX agent

arXiv:2604.12518v1 Announce Type: new Abstract: Multimodal sentiment analysis (MSA) integrates heterogeneous text, audio, and visual signals to infer human emotions. While recent approaches leverage c

researcharxiv-cs-cl
15 Apr 2026
Safety

Enhancing Agentic Textual Graph Retrieval with Synthetic Stepwise Supervision

DGX agent

arXiv:2510.03323v2 Announce Type: replace Abstract: Integrating textual graphs into Large Language Models (LLMs) is promising for complex graph-based QA. However, a key bottleneck is retrieving inform

safetyarxiv-cs-cl
15 Apr 2026
Applications

Evaluating Robustness of Large Language Models Against Multilingual Typographical Errors

DGX agent

arXiv:2510.09536v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly deployed in multilingual, real-world applications with user inputs -- naturally introducing typographi

applicationsarxiv-cs-cl
15 Apr 2026
Safety

EvoSpark: Endogenous Interactive Agent Societies for Unified Long-Horizon Narrative Evolution

DGX agent

arXiv:2604.12776v1 Announce Type: new Abstract: Realizing endogenous narrative evolution in LLM-based multi-agent systems is hindered by the inherent stochasticity of generative emergence. In particul

safetyarxiv-cs-cl
15 Apr 2026
Model Releases

FABLE: Fine-grained Fact Anchoring for Unstructured Model Editing

DGX agent

arXiv:2604.12559v1 Announce Type: new Abstract: Unstructured model editing aims to update models with real-world text, yet existing methods often memorize text holistically without reliable fine-grain

model-releasesarxiv-cs-cl
15 Apr 2026
Model Releases

From Imitation to Discrimination: Progressive Curriculum Learning for Robust Web Navigation

DGX agent

arXiv:2604.12666v1 Announce Type: cross Abstract: Text-based web agents offer computational efficiency for autonomous web navigation, yet developing robust agents remains challenging due to the noisy

model-releasesarxiv-cs-cl
15 Apr 2026
Safety

From Myopic Selection to Long-Horizon Awareness: Sequential LLM Routing for Multi-Turn Dialogue

DGX agent

arXiv:2604.12385v1 Announce Type: new Abstract: Multi-turn dialogue is the predominant form of interaction with large language models (LLMs). While LLM routing is effective in single-turn settings, ex

safetyarxiv-cs-cl
15 Apr 2026
Tutorials

Generating Effective CoT Traces for Mitigating Causal Hallucination

DGX agent

arXiv:2604.12748v1 Announce Type: new Abstract: Although large language models (LLMs) excel in complex reasoning tasks, they suffer from severe causal hallucination in event causality identification (

tutorialsarxiv-cs-cl
15 Apr 2026
Safety

GeoAlign: Geometric Feature Realignment for MLLM Spatial Reasoning

DGX agent

arXiv:2604.12630v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) have exhibited remarkable performance in various visual tasks, yet still struggle with spatial reasoning. Rec

safetyarxiv-cs-cl
15 Apr 2026
Local Ai

GigaCheck: Detecting LLM-generated Content via Object-Centric Span Localization

DGX agent

arXiv:2410.23728v3 Announce Type: replace Abstract: With the increasing quality and spread of LLM assistants, the amount of generated content is growing rapidly. In many cases and tasks, such texts ar

local-aiarxiv-cs-cl
15 Apr 2026
Research

GLeMM: A large-scale multilingual dataset for morphological research

DGX agent

arXiv:2604.12442v1 Announce Type: new Abstract: In derivational morphology, what mechanisms govern the variation in form-meaning relations between words? The answers to this type of questions are typi

researcharxiv-cs-cl
15 Apr 2026
Model Releases

GlotOCR Bench: OCR Models Still Struggle Beyond a Handful of Unicode Scripts

DGX agent

arXiv:2604.12978v1 Announce Type: new Abstract: Optical character recognition (OCR) has advanced rapidly with the rise of vision-language models, yet evaluation has remained concentrated on a small cl

model-releasesarxiv-cs-cl
15 Apr 2026
Applications

GRADE: Probing Knowledge Gaps in LLMs through Gradient Subspace Dynamics

DGX agent

arXiv:2604.02830v2 Announce Type: replace Abstract: Detecting whether a model's internal knowledge is sufficient to correctly answer a given question is a fundamental challenge in deploying responsibl

applicationsarxiv-cs-cl
15 Apr 2026
Safety

Gradient boundaries through confidence intervals for forced alignment estimates using model ensembles

DGX agent

arXiv:2506.01256v4 Announce Type: replace-cross Abstract: Forced alignment is a common tool to align audio with orthographic and phonetic transcriptions. Most forced alignment tools provide only point

safetyarxiv-cs-cl
15 Apr 2026
Safety

Graph-Based Chain-of-Thought Pruning for Reducing Redundant Reflections in Reasoning LLMs

DGX agent

arXiv:2604.05643v2 Announce Type: replace Abstract: Extending CoT through RL has been widely used to enhance the reasoning capabilities of LLMs. However, due to the sparsity of reward signals, it can

safetyarxiv-cs-cl
15 Apr 2026
Model Releases

Growing Pains: Extensible and Efficient LLM Benchmarking Via Fixed Parameter Calibration

DGX agent

arXiv:2604.12843v1 Announce Type: new Abstract: The rapid release of both language models and benchmarks makes it increasingly costly to evaluate every model on every dataset. In practice, models are

model-releasesarxiv-cs-cl
15 Apr 2026
Agents

Hear Both Sides: Efficient Multi-Agent Debate via Diversity-Aware Message Retention

DGX agent

arXiv:2603.20640v2 Announce Type: replace Abstract: Multi-Agent Debate has emerged as a promising framework for improving the reasoning quality of large language models through iterative inter-agent c

agentsarxiv-cs-cl
15 Apr 2026
Safety

InsightFlow: LLM-Driven Synthesis of Patient Narratives for Mental Health into Causal Models

DGX agent

arXiv:2604.12721v1 Announce Type: new Abstract: Clinical case formulation organizes patient symptoms and psychosocial factors into causal models, often using the 5P framework. However, constructing su

safetyarxiv-cs-cl
15 Apr 2026
Model Releases

KCLarity at SemEval-2026 Task 6: Encoder and Zero-Shot Approaches to Political Evasion Detection

DGX agent

arXiv:2603.06552v2 Announce Type: replace Abstract: This paper describes the KCLarity team's participation in CLARITY, a shared task at SemEval 2026 on classifying ambiguity and evasion techniques in

model-releasesarxiv-cs-cl
15 Apr 2026
Applications

Knowledge Is Not Static: Order-Aware Hypergraph RAG for Language Models

DGX agent

arXiv:2604.12185v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) enhances large language models by grounding outputs in retrieved knowledge. However, existing RAG methods including

applicationsarxiv-cs-cl
15 Apr 2026
Tutorials

KoCo: Conditioning Language Model Pre-training on Knowledge Coordinates

DGX agent

arXiv:2604.12397v1 Announce Type: new Abstract: Standard Large Language Model (LLM) pre-training typically treats corpora as flattened token sequences, often overlooking the real-world context that hu

tutorialsarxiv-cs-cl
15 Apr 2026
Research

Latent-Condensed Transformer for Efficient Long Context Modeling

DGX agent

arXiv:2604.12452v1 Announce Type: new Abstract: Large language models (LLMs) face significant challenges in processing long contexts due to the linear growth of the key-value (KV) cache and quadratic

researcharxiv-cs-cl
15 Apr 2026
Local Ai

Locate, Steer, and Improve: A Practical Survey of Actionable Mechanistic Interpretability in Large Language Models

DGX agent

arXiv:2601.14004v4 Announce Type: replace Abstract: Mechanistic Interpretability (MI) has emerged as a vital approach to demystify the opaque decision-making of Large Language Models (LLMs). However,

local-aiarxiv-cs-cl
15 Apr 2026
Research

LoSA: Locality Aware Sparse Attention for Block-Wise Diffusion Language Models

DGX agent

arXiv:2604.12056v1 Announce Type: new Abstract: Block-wise diffusion language models (DLMs) generate multiple tokens in any order, offering a promising alternative to the autoregressive decoding pipel

researcharxiv-cs-cl
15 Apr 2026
Local Ai

Masked by Consensus: Disentangling Privileged Knowledge in LLM Correctness

DGX agent

arXiv:2604.12373v1 Announce Type: new Abstract: Humans use introspection to evaluate their understanding through private internal states inaccessible to external observers. We investigate whether larg

local-aiarxiv-cs-cl
15 Apr 2026
Safety

Meet Dynamic Individual Preferences: Resolving Conflicting Human Value with Paired Fine-Tuning

DGX agent

arXiv:2604.12479v1 Announce Type: new Abstract: Recent advances in large language models (LLMs) have significantly improved the alignment of models with general human preferences. However, a major cha

safetyarxiv-cs-cl
15 Apr 2026
Research

MetFuse: Figurative Fusion between Metonymy and Metaphor

DGX agent

arXiv:2604.12919v1 Announce Type: new Abstract: Metonymy and metaphor often co-occur in natural language, yet computational work has studied them largely in isolation. We introduce a framework that tr

researcharxiv-cs-cl
15 Apr 2026
Model Releases

MoshiRAG: Asynchronous Knowledge Retrieval for Full-Duplex Speech Language Models

DGX agent

arXiv:2604.12928v1 Announce Type: new Abstract: Speech-to-speech language models have recently emerged to enhance the naturalness of conversational AI. In particular, full-duplex models are distinguis

model-releasesarxiv-cs-cl
15 Apr 2026
Research

Multilingual Multi-Label Emotion Classification at Scale with Synthetic Data

DGX agent

arXiv:2604.12633v1 Announce Type: new Abstract: Emotion classification in multilingual settings remains constrained by the scarcity of annotated data: existing corpora are predominantly English, singl

researcharxiv-cs-cl
15 Apr 2026
Agents

NaviRAG: Towards Active Knowledge Navigation for Retrieval-Augmented Generation

DGX agent

arXiv:2604.12766v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) typically relies on a flat retrieval paradigm that maps queries directly to static, isolated text segments. This ap

agentsarxiv-cs-cl
15 Apr 2026
Agents

OctoTools: An Agentic Framework with Extensible Tools for Complex Reasoning

DGX agent

arXiv:2502.11271v2 Announce Type: replace-cross Abstract: Solving complex reasoning tasks may involve visual understanding, domain knowledge retrieval, numerical calculation, and multi-step reasoning.

agentsarxiv-cs-cl
15 Apr 2026
Model Releases

Olmo 3

DGX agent

arXiv:2512.13961v2 Announce Type: replace Abstract: We introduce Olmo 3, a family of state-of-the-art, fully-open language models at the 7B and 32B parameter scales. Olmo 3 model construction targets

model-releasesarxiv-cs-cl
15 Apr 2026
Model Releases

ParetoBandit: Budget-Paced Adaptive Routing for Non-Stationary LLM Serving

DGX agent

arXiv:2604.00136v2 Announce Type: replace-cross Abstract: Multi-model LLM serving operates in a non-stationary, noisy environment: providers revise pricing, model quality can shift or regress without

model-releasesarxiv-cs-cl
15 Apr 2026
Safety

Perception-Aware Policy Optimization for Multimodal Reasoning

DGX agent

arXiv:2507.06448v5 Announce Type: replace Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has proven to be a highly effective strategy for endowing Large Language Models (LLMs) with ro

safetyarxiv-cs-cl
15 Apr 2026
Research

PILOT: Planning via Internalized Latent Optimization Trajectories for Large Language Models

DGX agent

arXiv:2601.19917v2 Announce Type: replace Abstract: Strategic planning is critical for multi-step reasoning, yet compact Large Language Models (LLMs) often lack the capacity to formulate global strate

researcharxiv-cs-cl
15 Apr 2026
Model Releases

PolicyLLM: Towards Excellent Comprehension of Public Policy for Large Language Models

DGX agent

arXiv:2604.12995v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly integrated into real-world decision-making, including in the domain of public policy. Yet, their ability t

model-releasesarxiv-cs-cl
15 Apr 2026
Model Releases

ReasonXL: Shifting LLM Reasoning Language Without Sacrificing Performance

DGX agent

arXiv:2604.12378v1 Announce Type: new Abstract: Despite advances in multilingual capabilities, most large language models (LLMs) remain English-centric in their training and, crucially, in their produ

model-releasesarxiv-cs-cl
15 Apr 2026
Research

Representing expertise accelerates learning from pedagogical interaction data

DGX agent

arXiv:2604.12195v1 Announce Type: new Abstract: Work in cognitive science and artificial intelligence has suggested that exposing learning agents to traces of interaction between multiple individuals

researcharxiv-cs-cl
15 Apr 2026
← Previous
1…148149150151152…160
Next →