AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
Model Releases

CGU-ILALab at FoodBench-QA 2026: Comparing Traditional and LLM-based Approaches for Recipe Nutrient Estimation

DGX agent

arXiv:2604.25774v1 Announce Type: new Abstract: Accurate nutrient estimation from unstructured recipe text is an important yet challenging problem in dietary monitoring, due to ambiguous ingredient te

model-releasesarxiv-cs-cl
29 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Cheaper, Better, Faster, Stronger: Robust Text-to-SQL without Chain-of-Thought or Fine-Tuning

DGX agent

arXiv:2505.14174v2 Announce Type: replace Abstract: LLMs are effective at code generation tasks like text-to-SQL, but is it worth the cost? Many state-of-the-art approaches use non-task-specific LLM t

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

Citation Failure: Definition, Analysis and Efficient Mitigation

DGX agent

arXiv:2510.20303v3 Announce Type: replace Abstract: Citations from LLM-based RAG systems are supposed to simplify response verification. However, this goal is undermined in cases of citation failure,

model-releasesarxiv-cs-cl
29 Apr 2026
Research

CodeOCR: On the Effectiveness of Vision Language Models in Code Understanding

DGX agent

arXiv:2602.01785v2 Announce Type: replace Abstract: Large Language Models (LLMs) have achieved remarkable success in source code understanding, yet as software systems grow in scale, computational eff

researcharxiv-cs-cl
29 Apr 2026
Agents

Cooperate to Compete: Strategic Coordination in Multi-Agent Conquest

DGX agent

arXiv:2604.25088v1 Announce Type: cross Abstract: Language Model (LM)-based agents remain largely untested in mixed-motive settings where agents must leverage short-term cooperation for long-term comp

agentsarxiv-cs-cl
29 Apr 2026
Safety

CORAL: Adaptive Retrieval Loop for Culturally-Aligned Multilingual RAG

DGX agent

arXiv:2604.25676v1 Announce Type: new Abstract: Multilingual retrieval-augmented generation (mRAG) is often implemented within a fixed retrieval space, typically via query or document translation or m

safetyarxiv-cs-cl
29 Apr 2026
Model Releases

CRAFT: Grounded Multi-Agent Coordination Under Partial Information

DGX agent

arXiv:2603.25268v2 Announce Type: replace Abstract: We introduce CRAFT, a multi-agent benchmark for evaluating pragmatic communication in large language models under strict partial information. In thi

model-releasesarxiv-cs-cl
29 Apr 2026
Safety

CroSearch-R1: Better Leveraging Cross-lingual Knowledge for Retrieval-Augmented Generation

DGX agent

arXiv:2604.25182v1 Announce Type: new Abstract: A multilingual collection may contain useful knowledge in other languages to supplement and correct the facts in the original language for Retrieval-Aug

safetyarxiv-cs-cl
29 Apr 2026
Model Releases

Cross-Lingual Jailbreak Detection via Semantic Codebooks

DGX agent

arXiv:2604.25716v1 Announce Type: new Abstract: Safety mechanisms for large language models (LLMs) remain predominantly English-centric, creating systematic vulnerabilities in multilingual deployment.

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

Cutscene Agent: An LLM Agent Framework for Automated 3D Cutscene Generation

DGX agent

arXiv:2604.25318v1 Announce Type: cross Abstract: Cutscenes are carefully choreographed cinematic sequences embedded in video games and interactive media, serving as the primary vehicle for narrative

model-releasesarxiv-cs-cl
29 Apr 2026
Agents

Diagnosis, Bad Planning & Reasoning. Treatment, SCOPE -- Planning for Hybrid Querying over Clinical Trial Data

DGX agent

arXiv:2604.25120v1 Announce Type: new Abstract: We study clinical trial table reasoning, where answers are not directly stored in visible cells but must be reasoned from semantic understanding through

agentsarxiv-cs-cl
29 Apr 2026
Research

DiffAdapt: Difficulty-Adaptive Reasoning for Token-Efficient LLM Inference

DGX agent

arXiv:2510.19669v4 Announce Type: replace Abstract: Recent reasoning Large Language Models (LLMs) demonstrate remarkable problem-solving abilities but often generate long thinking traces whose utility

researcharxiv-cs-cl
29 Apr 2026
Research

Do LLMs Capture Embodied Cognition and Cultural Variation? Cross-Linguistic Evidence from Demonstratives

DGX agent

arXiv:2604.25423v1 Announce Type: new Abstract: Do large language models (LLMs) truly acquire embodied cognition and cultural conventions from text? We introduce demonstratives, fundamental spatial ex

researcharxiv-cs-cl
29 Apr 2026
Model Releases

Doing More With Less: Revisiting the Effectiveness of LLM Pruning for Test-Time Scaling

DGX agent

arXiv:2604.25098v1 Announce Type: cross Abstract: While current Large Language Models (LLMs) exhibit remarkable reasoning capabilities through test-time compute scaling (TTS), their massive parameter

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

Dont Stop Early: Scalable Enterprise Deep Research with Controlled Information Flow and Evidence-Aware Termination

DGX agent

arXiv:2604.24978v1 Announce Type: new Abstract: Enterprise deep research often fails to produce decision-ready reports due to uneven information coverage, context explosion, and premature stopping. We

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

DRAGON: A Benchmark for Evidence-Grounded Visual Reasoning over Diagrams

DGX agent

arXiv:2604.25231v1 Announce Type: cross Abstract: Diagram question answering (DQA) requires models to interpret structured visual representations such as charts, maps, infographics, circuit schematics

model-releasesarxiv-cs-cl
29 Apr 2026
Local Ai

Dual-Track CoT: Budget-Aware Stepwise Guidance for Small LMs

DGX agent

arXiv:2604.25039v1 Announce Type: new Abstract: Large Language Models (LLMs) solve many reasoning tasks via chain-of-thought (CoT) prompting, but smaller models (about 7 to 8B parameters) still strugg

local-aiarxiv-cs-cl
29 Apr 2026
Model Releases

DV-World: Benchmarking Data Visualization Agents in Real-World Scenarios

DGX agent

arXiv:2604.25914v1 Announce Type: new Abstract: Real-world data visualization (DV) requires native environmental grounding, cross-platform evolution, and proactive intent alignment. Yet, existing benc

model-releasesarxiv-cs-cl
29 Apr 2026
Local Ai

Dynamic Decision Learning: Test-Time Evolution for Abnormality Grounding in Rare Diseases

DGX agent

arXiv:2604.24972v1 Announce Type: new Abstract: Clinical abnormality grounding for rare diseases is often hindered by data scarcity, making supervised fine-tuning impractical and single-pass inference

local-aiarxiv-cs-cl
29 Apr 2026
Research

Elderly-Contextual Data Augmentation via Speech Synthesis for Elderly ASR

DGX agent

arXiv:2604.24770v1 Announce Type: new Abstract: Despite recent progress in automatic speech recognition (ASR), elderly ASR (EASR) remains challenging due to limited training data and the distinct acou

researcharxiv-cs-cl
29 Apr 2026
Model Releases

Enhancing Financial Report Question-Answering: A Retrieval-Augmented Generation System with Reranking Analysis

DGX agent

arXiv:2603.16877v2 Announce Type: replace Abstract: Financial analysts face significant challenges extracting information from lengthy 10-K reports, which often exceed 100 pages. This paper presents a

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

Exploring Reasoning Reward Model for Agents

DGX agent

arXiv:2601.22154v2 Announce Type: replace-cross Abstract: Agentic Reinforcement Learning (Agentic RL) has achieved notable success in enabling agents to perform complex reasoning and tool use. However

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

Faithful Autoformalization via Roundtrip Verification and Repair

DGX agent

arXiv:2604.25031v1 Announce Type: new Abstract: When an LLM formalizes natural language, how do we know the output is faithful? We propose a roundtrip verification approach which does not require grou

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

Faithfulness-QA: A Counterfactual Entity Substitution Dataset for Training Context-Faithful RAG Models

DGX agent

arXiv:2604.25313v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) models frequently produce answers grounded in parametric memory rather than the retrieved context, undermining the

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

FAMA: Failure-Aware Meta-Agentic Framework for Open-Source LLMs in Interactive Tool Use Environments

DGX agent

arXiv:2604.25135v1 Announce Type: new Abstract: Large Language Models are being increasingly deployed as the decision-making core of autonomous agents capable of effecting change in external environme

model-releasesarxiv-cs-cl
29 Apr 2026
Safety

Frictive Policy Optimization for LLMs: Epistemic Intervention, Risk-Sensitive Control, and Reflective Alignment

DGX agent

arXiv:2604.25136v1 Announce Type: new Abstract: We propose Frictive Policy Optimization (FPO), a framework for learning language model policies that regulate not only what to say, but when and how to

safetyarxiv-cs-cl
29 Apr 2026
Research

From Ambiguity to Accuracy: The Transformative Effect of Coreference Resolution on Retrieval-Augmented Generation systems

DGX agent

arXiv:2507.07847v3 Announce Type: replace Abstract: Retrieval-Augmented Generation (RAG) has emerged as a crucial framework in natural language processing (NLP), improving factual consistency and redu

researcharxiv-cs-cl
29 Apr 2026
Research

From Chatbots to Confidants: A Cross-Cultural Study of LLM Adoption for Emotional Support

DGX agent

arXiv:2604.25525v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly used not only for instrumental tasks, but as always-available and non-judgmental confidants for emotional

researcharxiv-cs-cl
29 Apr 2026
Model Releases

From Local to Global: Revisiting Structured Pruning Paradigms for Large Language Models

DGX agent

arXiv:2510.18030v2 Announce Type: replace Abstract: Structured pruning is a practical approach to deploying large language models (LLMs) efficiently, as it yields compact, hardware-friendly architectu

model-releasesarxiv-cs-cl
29 Apr 2026
Research

From Syntax to Emotion: A Mechanistic Analysis of Emotion Inference in LLMs

DGX agent

arXiv:2604.25866v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used in emotionally sensitive human-AI applications, yet little is known about how emotion recognition is

researcharxiv-cs-cl
29 Apr 2026
Local Ai

From World-Gen to Quest-Line: A Dependency-Driven Prompt Pipeline for Coherent RPG Generation

DGX agent

arXiv:2604.25482v1 Announce Type: new Abstract: Large Language Models (LLMs) have shown strong potential for narrative generation, but their use in complex, multi-layered role-playing game (RPG) world

local-aiarxiv-cs-cl
29 Apr 2026
Model Releases

G-Loss: Graph-Guided Fine-Tuning of Language Models

DGX agent

arXiv:2604.25853v1 Announce Type: new Abstract: Traditional loss functions, including cross-entropy, contrastive, triplet, and su pervised contrastive losses, used for fine-tuning pre-trained language

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

GAIA-v2-LILT: Multilingual Adaptation of Agent Benchmark beyond Translation

DGX agent

arXiv:2604.24929v1 Announce Type: new Abstract: Agent benchmarks remain largely English-centric, while their multilingual versions are often built with machine translation (MT) and limited post-editin

model-releasesarxiv-cs-cl
29 Apr 2026
Safety

Generative AI Carries Non-Democratic Biases and Stereotypes: Representation of Women, Black Individuals, Age Groups, and People with Disability in AI-Generated Images across Occupations

DGX agent

arXiv:2409.13869v2 Announce Type: replace-cross Abstract: In this study, I investigate how generative artificial intelligence (AI) systems reproduce and reinforce societal biases, with a specific focu

safetyarxiv-cs-cl
29 Apr 2026
Safety

How RL Unlocks the Aha Moment in Geometric Interleaved Reasoning

DGX agent

arXiv:2603.01070v2 Announce Type: replace Abstract: Solving complex geometric problems inherently requires interleaved reasoning: a tight alternation between constructing diagrams and performing logic

safetyarxiv-cs-cl
29 Apr 2026
Model Releases

Images Amplify Misinformation Sharing in Vision-Language Models

DGX agent

arXiv:2505.13302v2 Announce Type: replace Abstract: As language and vision-language models (VLMs) become central to information access and online interaction, concerns grow about their potential to am

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

Improving LLM Predictions via Inter-Layer Structural Encoders

DGX agent

arXiv:2603.22665v2 Announce Type: replace Abstract: The standard practice in Large Language Models (LLMs) is to base predictions on final-layer representations. However, intermediate layers encode com

model-releasesarxiv-cs-cl
29 Apr 2026
Research

Independent-Component-Based Encoding Models of Brain Activity During Story Comprehension

DGX agent

arXiv:2604.24942v1 Announce Type: new Abstract: Encoding models provide a powerful framework for linking continuous stimulus features to neural activity; however, traditional voxelwise approaches are

researcharxiv-cs-cl
29 Apr 2026
Research

Intrinsic Mutual Information as a Modulator for Preference Optimization

DGX agent

arXiv:2604.24804v1 Announce Type: cross Abstract: Offline preference optimization methods, such as Direct Preference Optimization (DPO), offer significant advantages in aligning Large Language Models

researcharxiv-cs-cl
29 Apr 2026
Research

Is Large Language Model Performance on Reasoning Tasks Impacted by Different Ways Questions Are Asked?

DGX agent

arXiv:2507.15707v2 Announce Type: replace Abstract: Large Language Models (LLMs) have been evaluated using diverse question types, e.g., multiple-choice, true/false, and short/long answers. This study

researcharxiv-cs-cl
29 Apr 2026
Model Releases

Is This Just Fantasy? Language Model Representations Reflect Human Judgments of Event Plausibility

DGX agent

arXiv:2507.12553v3 Announce Type: replace Abstract: Language models (LMs) are used for a diverse range of tasks, from question answering to writing fantastical stories. In order to reliably accomplish

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

jina-embeddings-v5-text: Task-Targeted Embedding Distillation

DGX agent

arXiv:2602.15547v2 Announce Type: replace Abstract: Text embedding models are widely used for semantic similarity tasks, including information retrieval, clustering, and classification. General-purpos

model-releasesarxiv-cs-cl
29 Apr 2026
Research

Korean aegyo speech shows systematic F1 increase to signal childlike qualities

DGX agent

arXiv:2604.25133v1 Announce Type: new Abstract: Korean aegyo is a socially recognized childlike speaking style used predominantly in romantic interactions among adults. This study examined vowel space

researcharxiv-cs-cl
29 Apr 2026
Research

Language corpora for the Dutch medical domain

DGX agent

arXiv:2604.25374v1 Announce Type: new Abstract: extbf{Background:} Dutch medical corpora are scarce, limiting NLP development. extbf{Methods:} We translated English datasets, identified medical text i

researcharxiv-cs-cl
29 Apr 2026
Research

Large Language Models Are Effective Human Annotation Assistants, But Not Good Independent Annotators

DGX agent

arXiv:2503.06778v3 Announce Type: replace Abstract: Event annotation is important for identifying market changes, monitoring breaking news, and understanding sociological trends. Although expert annot

researcharxiv-cs-cl
29 Apr 2026
Model Releases

Large Language Models Explore by Latent Distilling

DGX agent

arXiv:2604.24927v1 Announce Type: new Abstract: Generating diverse responses is crucial for test-time scaling of large language models (LLMs), yet standard stochastic sampling mostly yields surface-le

model-releasesarxiv-cs-cl
29 Apr 2026
Safety

Learning-Based Automated Adversarial Red-Teaming for Robustness Evaluation of Large Language Models

DGX agent

arXiv:2512.20677v4 Announce Type: replace-cross Abstract: The increasing deployment of large language models (LLMs) in safety-critical applications raises fundamental challenges in systematically eval

safetyarxiv-cs-cl
29 Apr 2026
Safety

Learning from Medical Entity Trees: An Entity-Centric Medical Data Engineering Framework for MLLMs

DGX agent

arXiv:2604.25296v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have shown transformative potential in medical applications, yet their performance is hindered by conventional

safetyarxiv-cs-cl
29 Apr 2026
← Previous
1…118119120121122…161
Next →