AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
Research

DRP: Distilled Reasoning Pruning with Skill-aware Step Decomposition for Efficient Large Reasoning Models

DGX agent

arXiv:2505.13975v4 Announce Type: replace Abstract: While Large Reasoning Models (LRMs) have demonstrated success in complex reasoning tasks through long chain-of-thought (CoT) reasoning, their infere

researcharxiv-cs-cl
28 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Applications

DualGuard: Dual-stream Large Language Model Watermarking Defense against Paraphrase and Spoofing Attack

DGX agent

arXiv:2512.16182v2 Announce Type: replace-cross Abstract: With the rapid development of cloud-based services, large language models have become increasingly accessible through various web platforms. H

applicationsarxiv-cs-cl
28 Apr 2026
Safety

Efficient Rationale-based Retrieval: On-policy Distillation from Generative Rerankers based on JEPA

DGX agent

arXiv:2604.23336v1 Announce Type: cross Abstract: Unlike traditional fact-based retrieval, rationale-based retrieval typically necessitates cross-encoding of query-document pairs using large language

safetyarxiv-cs-cl
28 Apr 2026
Model Releases

EgoDyn-Bench: Evaluating Ego-Motion Understanding in Vision-Centric Foundation Models for Autonomous Driving

DGX agent

arXiv:2604.22851v1 Announce Type: cross Abstract: While Vision-Language Models (VLMs) have advanced highlevel reasoning in autonomous driving, their ability to ground this reasoning in the underlying

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

Evaluating Large Language Models on Computer Science University Exams in Data Structures

DGX agent

arXiv:2604.23347v1 Announce Type: new Abstract: We present a comprehensive evaluation of Large Language Models (LLMs) on Computer Science (CS) Data Structure examination questions. Our work introduces

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

Evaluating Temporal Consistency in Multi-Turn Language Models

DGX agent

arXiv:2604.23051v1 Announce Type: new Abstract: Language models are increasingly deployed in interactive settings where users reason about facts over time rather than in isolation. In such scenarios,

model-releasesarxiv-cs-cl
28 Apr 2026
Research

Evaluation Framework for Highlight Explanations of Context Utilisation in Language Models

DGX agent

arXiv:2510.02629v3 Announce Type: replace Abstract: Context utilisation, the ability of Language Models (LMs) to incorporate relevant information from the provided context when generating responses, r

researcharxiv-cs-cl
28 Apr 2026
Research

Evaluation of Pose Estimation Systems for Sign Language Translation

DGX agent

arXiv:2604.24609v1 Announce Type: new Abstract: Many sign language translation (SLT) systems operate on pose sequences instead of raw video to reduce input dimensionality, improve portability, and par

researcharxiv-cs-cl
28 Apr 2026
Model Releases

Evolve: A Persistent Knowledge Lifecycle for Small Language Models

DGX agent

arXiv:2604.23424v1 Announce Type: cross Abstract: Evolve pairs a small local language model with a persistent, teacher-compiled knowledge store -- refined through sleep consolidation and usage-driven

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

EXCEEDS: Extracting Complex Events via Nugget-based Grid Modeling in Scientific Domain

DGX agent

arXiv:2406.14075v2 Announce Type: replace Abstract: It is crucial to understand a specific domain by events. Extensive event extraction research has been conducted in many domains such as news, financ

model-releasesarxiv-cs-cl
28 Apr 2026
Research

Explaining Sources of Uncertainty in Automated Fact-Checking

DGX agent

arXiv:2505.17855v2 Announce Type: replace Abstract: Understanding sources of a model's uncertainty regarding its predictions is crucial for effective human-AI collaboration. Prior work proposes using

researcharxiv-cs-cl
28 Apr 2026
Local Ai

Factual and Edit-Sensitive Graph-to-Sequence Generation via Graph-Aware Adaptive Noising

DGX agent

arXiv:2604.24104v1 Announce Type: new Abstract: Fine-tuned autoregressive models for graph-to-sequence generation (G2S) often struggle with factual grounding and edit sensitivity. To tackle these issu

local-aiarxiv-cs-cl
28 Apr 2026
Tutorials

Fine-tuning vs. In-context Learning in Large Language Models: A Formal Language Learning Perspective

DGX agent

arXiv:2604.23267v1 Announce Type: new Abstract: Large language models (LLMs) operate in two fundamental learning modes - fine-tuning (FT) and in-context learning (ICL) - raising key questions about wh

tutorialsarxiv-cs-cl
28 Apr 2026
Agents

Food4All: A Multi-Agent Framework for Real-time Free Food Discovery with Integrated Nutritional Metadata

DGX agent

arXiv:2510.18289v2 Announce Type: replace Abstract: Food insecurity remains a persistent public health emergency in the United States, tightly interwoven with chronic disease, mental illness, and opio

agentsarxiv-cs-cl
28 Apr 2026
Model Releases

For-Value: Efficient Forward-Only Data Valuation for finetuning LLMs and VLMs

DGX agent

arXiv:2508.10180v3 Announce Type: replace Abstract: Data valuation is essential for enhancing the transparency and accountability of large language models (LLMs) and vision-language models (VLMs). How

model-releasesarxiv-cs-cl
28 Apr 2026
Research

Gated Tree Cross-Attention for Checkpoint-Compatible Syntax Injection in Decoder-Only LLMs

DGX agent

arXiv:2602.15846v2 Announce Type: replace Abstract: Decoder-only large language models achieve strong broad performance but are brittle to minor grammatical perturbations, undermining reliability for

researcharxiv-cs-cl
28 Apr 2026
Model Releases

Generating Place-Based Compromises Between Two Points of View

DGX agent

arXiv:2604.24536v1 Announce Type: new Abstract: Large Language Models (LLMs) excel academically but struggle with social intelligence tasks, such as creating good compromises. In this paper, we presen

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

GraphPlanner: Graph Memory-Augmented Agentic Routing for Multi-Agent LLMs

DGX agent

arXiv:2604.23626v1 Announce Type: new Abstract: LLM routing has achieved promising results in integrating the strengths of diverse models while balancing efficiency and performance. However, to suppor

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

HalalBench: A Multilingual OCR Benchmark for Food Packaging Ingredient Extraction

DGX agent

arXiv:2604.22754v1 Announce Type: cross Abstract: No standardized benchmark exists for evaluating OCR on food packaging, despite its critical role in automated halal food verification. Existing benchm

model-releasesarxiv-cs-cl
28 Apr 2026
Research

HeadRouter: Dynamic Head-Weight Routing for Task-Adaptive Audio Token Pruning in Large Audio Language Models

DGX agent

arXiv:2604.23717v1 Announce Type: cross Abstract: Recent large audio language models (LALMs) demonstrate remarkable capabilities in processing extended multi-modal sequences, yet incur high inference

researcharxiv-cs-cl
28 Apr 2026
Local Ai

Hidden States Know Where Reasoning Diverges: Credit Assignment via Span-Level Wasserstein Distance

DGX agent

arXiv:2604.23318v1 Announce Type: new Abstract: Group Relative Policy Optimization (GRPO) performs coarse-grained credit assignment in reinforcement learning with verifiable rewards (RLVR) by assignin

local-aiarxiv-cs-cl
28 Apr 2026
Model Releases

How Sensitive Are Safety Benchmarks to Judge Configuration Choices?

DGX agent

arXiv:2604.24074v1 Announce Type: new Abstract: Safety benchmarks such as HarmBench rely on LLM judges to classify model responses as harmful or safe, yet the judge configuration, namely the combinati

model-releasesarxiv-cs-cl
28 Apr 2026
Applications

Implicit Framing in Obstetric Counseling Notes: A Grounded LLM Pipeline on a VBAC-Eligible Cohort

DGX agent

arXiv:2604.23059v1 Announce Type: new Abstract: Clinical framing -- the linguistic manner in which clinical information is presented -- can influence patient understanding and decision-making, with im

applicationsarxiv-cs-cl
28 Apr 2026
Research

In-depth Analysis of Graph-based RAG in a Unified Framework

DGX agent

arXiv:2503.04338v2 Announce Type: replace-cross Abstract: Graph-based Retrieval-Augmented Generation (RAG) has proven effective in integrating external knowledge into large language models (LLMs), imp

researcharxiv-cs-cl
28 Apr 2026
Safety

In-Sync: Adaptation of Speech Aware Large Language Models for ASR with Word Level Timestamp Predictions

DGX agent

arXiv:2604.22817v1 Announce Type: cross Abstract: Recent advances in speech-aware language models have coupled strong acoustic encoders with large language models, enabling systems that move beyond tr

safetyarxiv-cs-cl
28 Apr 2026
Research

Indirect Question Answering in English, German and Bavarian: A Challenging Task for High- and Low-Resource Languages Alike

DGX agent

arXiv:2603.15130v2 Announce Type: replace Abstract: Indirectness is a common feature of daily communication, yet is underexplored in NLP research for both low-resource as well as high-resource languag

researcharxiv-cs-cl
28 Apr 2026
Tutorials

Investigating the Representation of Backchannels and Fillers in Fine-tuned Language Models

DGX agent

arXiv:2509.20237v2 Announce Type: replace Abstract: Backchannels and fillers are important linguistic expressions in dialogue, but often treated as 'noise' to be bypassed in modern transformer-based l

tutorialsarxiv-cs-cl
28 Apr 2026
Safety

IRIS: Interleaved Reinforcement with Incremental Staged Curriculum for Cross-Lingual Mathematical Reasoning

DGX agent

arXiv:2604.24114v1 Announce Type: new Abstract: Curriculum learning helps language models tackle complex reasoning by gradually increasing task difficulty. However, it often fails to generate consiste

safetyarxiv-cs-cl
28 Apr 2026
Model Releases

JudgeSense: A Benchmark for Prompt Sensitivity in LLM-as-a-Judge Systems

DGX agent

arXiv:2604.23478v1 Announce Type: new Abstract: Large language models are increasingly deployed as automated judges for evaluating other models, yet the stability of their verdicts under semantically

model-releasesarxiv-cs-cl
28 Apr 2026
Research

Knowledge Vector of Logical Reasoning in Large Language Models

DGX agent

arXiv:2604.23877v1 Announce Type: new Abstract: Logical reasoning serve as a central capability in LLMs and includes three main forms: deductive, inductive, and abductive reasoning. In this work, we s

researcharxiv-cs-cl
28 Apr 2026
Research

Large language model-enabled automated data extraction for concrete materials informatics

DGX agent

arXiv:2604.22938v1 Announce Type: cross Abstract: The promise of data-driven materials discovery remains constrained by the scarcity of large, high-quality, and accessible experimental datasets. Here,

researcharxiv-cs-cl
28 Apr 2026
Research

Learning Evidence of Depression Symptoms via Prompt Induction

DGX agent

arXiv:2604.24376v1 Announce Type: new Abstract: Depression places substantial pressure on mental health services, and many people describe their experiences outside clinical settings in high-volume us

researcharxiv-cs-cl
28 Apr 2026
Safety

Learning Selective LLM Autonomy from Copilot Feedback in Enterprise Customer Support Workflows

DGX agent

arXiv:2604.23855v1 Announce Type: new Abstract: We present a deployed system that automates end-to-end customer support workflows inside an enterprise Business Process Management (BPM) platform. The a

safetyarxiv-cs-cl
28 Apr 2026
Applications

LegalDrill: Diagnosis-Driven Synthesis for Legal Reasoning in Small Language Models

DGX agent

arXiv:2604.23809v1 Announce Type: new Abstract: Small language models (SLMs) are promising for real-world deployment due to their efficiency and low operational cost. However, their limited capacity s

applicationsarxiv-cs-cl
28 Apr 2026
Safety

LinguDistill: Recovering Linguistic Ability in Vision-Language Models via Selective Cross-Modal Distillation

DGX agent

arXiv:2604.00829v3 Announce Type: replace-cross Abstract: Adapting pretrained language models (LMs) into vision-language models (VLMs) can degrade their native linguistic capability due to representat

safetyarxiv-cs-cl
28 Apr 2026
Model Releases

Long-Context Aware Upcycling: A New Frontier for Hybrid LLM Scaling

DGX agent

arXiv:2604.24715v1 Announce Type: new Abstract: Hybrid sequence models that combine efficient Transformer components with linear sequence modeling blocks are a promising alternative to pure Transforme

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

LongFlow: Efficient KV Cache Compression for Reasoning Models

DGX agent

arXiv:2603.11504v2 Announce Type: replace-cross Abstract: Recent reasoning models such as OpenAI-o1 and DeepSeek-R1 have shown strong performance on complex tasks including mathematical reasoning and

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

Looking for the Bottleneck in Fine-grained Temporal Relation Classification

DGX agent

arXiv:2604.24620v1 Announce Type: new Abstract: Temporal relation classification is the task of determining the temporal relation between pairs of temporal entities in a text. Despite recent advanceme

model-releasesarxiv-cs-cl
28 Apr 2026
Research

Making Dialogue Grounding Data Rich: A Three-Tier Data Synthesis Framework for Generalized Referring Expression Comprehension

DGX agent

arXiv:2512.02791v2 Announce Type: replace Abstract: Dialogue-Based Generalized Referring Expression Comprehension (GREC) requires models to ground the expression and unlimited targets in complex visua

researcharxiv-cs-cl
28 Apr 2026
Research

Measuring Temporal Linguistic Emergence in Diffusion Language Models

DGX agent

arXiv:2604.23235v1 Announce Type: new Abstract: Diffusion language models expose an explicit denoising trajectory, making it possible to ask when different kinds of information become measurable durin

researcharxiv-cs-cl
28 Apr 2026
Model Releases

MEG-RAG: Quantifying Multi-modal Evidence Grounding for Evidence Selection in RAG

DGX agent

arXiv:2604.24564v1 Announce Type: new Abstract: Multimodal Retrieval-Augmented Generation (MRAG) addresses key limitations of Multimodal Large Language Models (MLLMs), such as hallucination and outdat

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

Mind the Gap: Evaluating Model- and Agentic-Level Vulnerabilities in LLMs with Action Graphs

DGX agent

arXiv:2509.04802v3 Announce Type: replace Abstract: As large language models increasingly deployed into agentic systems, existing methods face critical gaps in observing, assessing, and mitigating dep

model-releasesarxiv-cs-cl
28 Apr 2026
Safety

MIPIC: Matryoshka Representation Learning via Self-Distilled Intra-Relational and Progressive Information Chaining

DGX agent

arXiv:2604.24374v1 Announce Type: new Abstract: Representation learning is fundamental to NLP, but building embeddings that work well at different computational budgets is challenging. Matryoshka Repr

safetyarxiv-cs-cl
28 Apr 2026
Research

Multimodal QUD: Inquisitive Questions from Scientific Figures

DGX agent

arXiv:2604.23733v1 Announce Type: new Abstract: Asking inquisitive questions while reading, and looking for their answers, is an important part in human discourse comprehension, curiosity, and creativ

researcharxiv-cs-cl
28 Apr 2026
Research

Neural Grammatical Error Correction for Romanian

DGX agent

arXiv:2604.23627v1 Announce Type: new Abstract: Resources for Grammatical Error Correction (GEC) in non-English languages are scarce, while available spellcheckers in these languages are mostly limite

researcharxiv-cs-cl
28 Apr 2026
Model Releases

OLaPh: Optimal Language Phonemizer

DGX agent

arXiv:2509.20086v2 Announce Type: replace Abstract: Phonemization is a critical component in text-to-speech synthesis. Traditional approaches rely on deterministic transformations and lexica, while ne

model-releasesarxiv-cs-cl
28 Apr 2026
Research

On Emergent Social World Models -- Evidence for Functional Integration of Theory of Mind and Pragmatic Reasoning in Language Models

DGX agent

arXiv:2602.10298v2 Announce Type: replace Abstract: This paper investigates whether LMs recruit shared computational mechanisms for general Theory of Mind (ToM) and language-specific pragmatic reasoni

researcharxiv-cs-cl
28 Apr 2026
Applications

One Size Fits None: Heuristic Collapse in LLM Investment Advice

DGX agent

arXiv:2604.23837v1 Announce Type: new Abstract: Large language models are increasingly deployed as advisors in high-stakes domains -- answering medical questions, interpreting legal documents, recomme

applicationsarxiv-cs-cl
28 Apr 2026
← Previous
1…121122123124125…161
Next →