AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlog
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,759 results
Model Releases

Learning Query-Aware Budget-Tier Routing for Runtime Agent Memory

DGX agent

arXiv:2602.06025v2 Announce Type: replace Abstract: Memory is increasingly central to Large Language Model (LLM) agents operating beyond a single context window, yet most existing systems rely on offl

model-releasesarxiv-cs-cl
21 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Research

Leveraging Large Language Models for Sentiment Analysis: Multi-Modal Analysis of Decentraland's MANA Token

DGX agent

arXiv:2605.20192v1 Announce Type: new Abstract: Decentraland, a decentralized virtual reality platform operating within the expanding Metaverse ecosystem, utilizes its native MANA token to facilitate

researcharxiv-cs-cl
21 May 2026
Model Releases

Leveraging LLMs for Grammar Adaptation: A Study on Metamodel-Grammar Co-Evolution

DGX agent

arXiv:2605.21465v1 Announce Type: new Abstract: In model-driven engineering, metamodel evolution leads to the need to adapt corresponding grammars to maintain consistency, which typically requires ted

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

LLMs on the Line: Data Determines Loss-to-Loss Scaling Laws

DGX agent

arXiv:2502.12120v3 Announce Type: replace-cross Abstract: Scaling laws guide the development of large language models (LLMs) by offering estimates for the optimal balance of model size, tokens, and co

model-releasesarxiv-cs-cl
21 May 2026
Local Ai

LoCar: Localization-Aware Evaluation of In-Vehicle Assistants through Fine-Grained Sociolinguistic Control

DGX agent

arXiv:2605.21086v1 Announce Type: new Abstract: While Large Language Models (LLMs) are increasingly integrated into in-vehicle conversational systems, identifying the optimal model remains challenging

local-aiarxiv-cs-cl
21 May 2026
Research

Long-Context Reasoning Through Proxy-Based Chain-of-Thought Tuning

DGX agent

arXiv:2605.20201v1 Announce Type: new Abstract: Recent large language models support inputs of up to 10 million tokens, yet they perform poorly on long-context tasks that require complex reasoning. Su

researcharxiv-cs-cl
21 May 2026
Research

Manga109-v2026: Revisiting Manga109 Annotations for Modern Manga Understanding

DGX agent

arXiv:2605.21182v1 Announce Type: new Abstract: Manga is a culturally distinctive multimodal medium and one of the most influential forms of Japanese popular culture. As AI systems increasingly target

researcharxiv-cs-cl
21 May 2026
Agents

MASFactory: A Graph-centric Framework for Orchestrating LLM-Based Multi-Agent Systems with Vibe Graphing

DGX agent

arXiv:2603.06007v2 Announce Type: replace Abstract: Large language model-based (LLM-based) multi-agent systems (MAS) are increasingly used to extend agentic problem solving via role specialization and

agentsarxiv-cs-cl
21 May 2026
Applications

Measuring and mitigating overreliance to build human-compatible AI

DGX agent

arXiv:2509.08010v2 Announce Type: replace-cross Abstract: Large language models (LLMs) distinguish themselves from previous technologies by functioning as collaborative ``thought partners,'' capable o

applicationsarxiv-cs-cl
21 May 2026
Model Releases

Mechanics of Bias and Reasoning: Interpreting the Impact of Chain-of-Thought Prompting on Gender Bias in LLMs

DGX agent

arXiv:2605.20410v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed in socially sensitive settings despite substantial documentation that they encode gender biases.

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

MedicalBench: Evaluating Large Language Models Toward Improved Medical Concept Extraction

DGX agent

arXiv:2605.20197v1 Announce Type: new Abstract: Medical concept extraction from electronic health records underpins many downstream applications, yet remains challenging because medically meaningful c

model-releasesarxiv-cs-cl
21 May 2026
Agents

Mem-pi: Adaptive Memory through Learning When and What to Generate

DGX agent

arXiv:2605.21463v1 Announce Type: new Abstract: We present Mem-pi, a framework for adaptive memory in large language model (LLM) agents, where useful guidance is generated on demand rather than retrie

agentsarxiv-cs-cl
21 May 2026
Model Releases

MemGym: a Long-Horizon Memory Environment for LLM Agents

DGX agent

arXiv:2605.20833v1 Announce Type: new Abstract: Memory is a central capability for LLM agents operating across long-horizon tasks. Existing memory benchmarks predominantly evaluate retention of person

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

Memory Grafting: Scaling Language Model Pre-training via Offline Conditional Memory

DGX agent

arXiv:2605.20948v1 Announce Type: new Abstract: Scaling conditional memory offers a promising way to increase language-model capacity, but existing methods such as Engram learn large memory tables fro

model-releasesarxiv-cs-cl
21 May 2026
Research

Metaphors in Literary Post-Editing: Opening Pandora's Box?

DGX agent

arXiv:2605.21178v1 Announce Type: new Abstract: This paper investigates how post-editors of literary texts react and respond to the way metaphors have been translated by Neu ral Machine Translation (N

researcharxiv-cs-cl
21 May 2026
Agents

Mix-Quant: Quantized Prefilling, Precise Decoding for Agentic LLMs

DGX agent

arXiv:2605.20315v1 Announce Type: new Abstract: LLM agents have recently emerged as a powerful paradigm for solving complex tasks through planning, tool use, memory retrieval, and multi-step interacti

agentsarxiv-cs-cl
21 May 2026
Research

Most Transformer Modifications Still Do Not Transfer at 1-3B: A 2020-2026 Update to Narang et al. (2021) with Downstream Evaluation and a Noise Floor

DGX agent

arXiv:2605.20798v1 Announce Type: cross Abstract: Narang et al. (2021) evaluated 40+ Transformer modifications at T5-base scale and concluded that most did not transfer. Five years later, the typical

researcharxiv-cs-cl
21 May 2026
Model Releases

MTR-Suite: A Framework for Evaluating and Synthesizing Conversational Retrieval Benchmarks

DGX agent

arXiv:2605.20729v1 Announce Type: new Abstract: Accurate evaluation of conversational retrieval is pivotal for advancing Retrieval-Augmented Generation (RAG) systems. However, existing conversational

model-releasesarxiv-cs-cl
21 May 2026
Agents

Multi-agent Collaboration with State Management

DGX agent

arXiv:2605.20563v1 Announce Type: cross Abstract: Recent advances in multi-agent systems have shown great potential for solving complex tasks. However, when multiple agents edit a shared codebase conc

agentsarxiv-cs-cl
21 May 2026
Model Releases

NeuroQA: A Large-Scale Image-Grounded Benchmark for 3D Brain MRI Understanding

DGX agent

arXiv:2605.20525v1 Announce Type: cross Abstract: We present NeuroQA, a large-scale benchmark for visual question answering in 3D brain magnetic resonance imaging (MRI), with 56,953 QA pairs from 12,9

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

On the limits and opportunities of AI reviewers: Reviewing the reviews of Nature-family papers with 45 expert scientists

DGX agent

arXiv:2605.20668v1 Announce Type: new Abstract: With the advancement of AI capabilities, AI reviewers are beginning to be deployed in scientific peer review, yet their capability and credibility remai

model-releasesarxiv-cs-cl
21 May 2026
Research

Optimal Query Allocation in Extractive QA with LLMs: A Learning-to-Defer Framework with Theoretical Guarantees

DGX agent

arXiv:2410.15761v4 Announce Type: replace Abstract: Large Language Models excel in generative tasks but exhibit inefficiencies in structured text selection, particularly in extractive question answeri

researcharxiv-cs-cl
21 May 2026
Safety

Parallel LLM Reasoning for Bias-Resilient, Robust Conceptual Abstraction

DGX agent

arXiv:2605.20194v1 Announce Type: new Abstract: Large language models (LLMs) have been increasingly used to analyze text. However, they are often plagued with contextual reasoning limitations when ana

safetyarxiv-cs-cl
21 May 2026
Research

Playing Devil's Advocate: Off-the-Shelf Persona Vectors Rival Targeted Steering for Sycophancy

DGX agent

arXiv:2605.21006v1 Announce Type: cross Abstract: We study the effect of different persona on extbf{sycophancy}: model's agreement with users even when the user is incorrect. The standard mitigation,

researcharxiv-cs-cl
21 May 2026
Model Releases

Post-Hoc Understanding of Metaphor Processing in Decoder-Only Language Models via Conditional Scale Entropy

DGX agent

arXiv:2605.21391v1 Announce Type: new Abstract: Metaphor requires a language model to resolve a token whose contextual meaning diverges from its basic literal sense. Understanding how transformer mode

model-releasesarxiv-cs-cl
21 May 2026
Tutorials

Pseudo-Siamese Network for Planning in Target-Oriented Proactive Dialogues

DGX agent

arXiv:2605.20195v1 Announce Type: new Abstract: A target-oriented proactive dialogue system is designed to steer conversations toward predefined targets while actively providing suggestions. The core

tutorialsarxiv-cs-cl
21 May 2026
Hardware

PulseCol: Periodically Refreshed Column-Sparse Attention for Accelerating Diffusion Language Models

DGX agent

arXiv:2605.20813v1 Announce Type: new Abstract: Inference in diffusion large language models (dLLMs) is computationally expensive, as full self-attention must be repeatedly executed at each step of th

hardwarearxiv-cs-cl
21 May 2026
Tutorials

Puzzled By ChatGPT? No more! A Jigsaw Puzzle to Promote AI Literacy and Awareness

DGX agent

arXiv:2605.20404v1 Announce Type: new Abstract: The rapid adoption of Generative AI, including LLM-based chatbots like ChatGPT, has highlighted the need for accessible ways to support public understan

tutorialsarxiv-cs-cl
21 May 2026
Research

Quantifying the cross-linguistic effects of syncretism on agreement attraction

DGX agent

arXiv:2605.21403v1 Announce Type: new Abstract: Agreement attraction errors, in which a verb erroneously agrees with an intervening noun rather than its grammatical head, are amplified by morphologica

researcharxiv-cs-cl
21 May 2026
Model Releases

Refining and Reusing Annotation Guidelines for LLM Annotation

DGX agent

arXiv:2605.20809v1 Announce Type: new Abstract: While Large Language Models (LLMs) demonstrate remarkable performance on zero-shot annotation tasks, they often struggle with the specialized convention

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

Reinforcing Human Behavior Simulation via Verbal Feedback

DGX agent

arXiv:2605.20506v1 Announce Type: cross Abstract: Humans learn social norms and behaviors from verbal feedback (e.g., a parent saying 'that was rude' or a friend explaining 'here's why that hurt'). Ye

model-releasesarxiv-cs-cl
21 May 2026
Research

Reliable Automated Triage in Spanish Clinical Notes: A Hybrid Framework for Risk-Aware HIV Suspicion Identification

DGX agent

arXiv:2605.21256v1 Announce Type: new Abstract: Standard clinical Natural Language Processing (NLP) benchmarks often yield inflated metrics by forcing deterministic classification on ambiguous instanc

researcharxiv-cs-cl
21 May 2026
Agents

Retrieval-Augmented Code Generation: A Survey with Focus on Repository-Level Approaches

DGX agent

arXiv:2510.04905v3 Announce Type: replace-cross Abstract: Recent advances in large language models (LLMs) have significantly improved automated code generation. While existing approaches have achieved

agentsarxiv-cs-cl
21 May 2026
Model Releases

Retrieval-Augmented Long-Context Translation for Cultural Image Captioning: Gators submission for AmericasNLP 2026 shared task

DGX agent

arXiv:2605.20626v1 Announce Type: new Abstract: We present the University of Florida Gators submission to the AmericasNLP 2026 shared task on cultural image captioning for Indigenous languages. Our tw

model-releasesarxiv-cs-cl
21 May 2026
Research

Retrospective Sparse Attention for Efficient Long-Context Generation

DGX agent

arXiv:2508.09001v2 Announce Type: replace Abstract: Large Language Models (LLMs) are increasingly deployed in long-context tasks such as reasoning, code generation, and multi-turn dialogue. However, i

researcharxiv-cs-cl
21 May 2026
Safety

SCRIBE: Diagnostic Evaluation and Rich Transcription Models for Indic ASR

DGX agent

arXiv:2605.20712v1 Announce Type: new Abstract: Automatic speech recognition replaces typing only when correction costs less than manual entry, a threshold determined by error types, not counts: fixin

safetyarxiv-cs-cl
21 May 2026
Research

Self-Training Doesn't Flatten Language -- It Restructures It: Surface Markers Amplify While Deep Syntax Dies

DGX agent

arXiv:2605.20602v1 Announce Type: new Abstract: Successive self-training on a language model's own outputs is widely characterized as a process of flattening: diversity drops, distributions narrow, an

researcharxiv-cs-cl
21 May 2026
Model Releases

SHINE: A Scalable In-Context Hypernetwork for Mapping Context to LoRA in a Single Pass

DGX agent

arXiv:2602.06358v2 Announce Type: replace Abstract: We propose SHINE (Scalable Hyper In-context NEtwork), a scalable hypernetwork that can map diverse meaningful contexts into high-quality LoRA adapte

model-releasesarxiv-cs-cl
21 May 2026
Safety

Shiny Stories, Hidden Struggles: Investigating the Representation of Disability Through the Lens of LLMs

DGX agent

arXiv:2605.20191v1 Announce Type: new Abstract: Modern Large Language Models (LLMs) have recently attracted much attention for their ability to simulate human behavior and generate text that reflects

safetyarxiv-cs-cl
21 May 2026
Research

Single-Pass, Depth-Selective Reading for Multi-Aspect Sentiment Analysis

DGX agent

arXiv:2605.20998v1 Announce Type: new Abstract: Aspect-Term Sentiment Analysis (ATSA) in multi-aspect sentences faces a fundamental tradeoff between efficiency and expressiveness. Existing models eith

researcharxiv-cs-cl
21 May 2026
Research

Smarter edits? Post-editing with error highlights and translation suggestions

DGX agent

arXiv:2605.21135v1 Announce Type: new Abstract: As MT quality increases, interest in enhanced post-editing features such as QE-derived error highlights is growing, yet evidence for their usefulness re

researcharxiv-cs-cl
21 May 2026
Model Releases

SMoA: Spectrum Modulation Adapter for Parameter-Efficient Fine-Tuning

DGX agent

arXiv:2605.21147v1 Announce Type: cross Abstract: As the number of model parameters increases, parameter-efficient fine-tuning (PEFT) has become the go-to choice for tailoring pre-trained large langua

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

SpecBench: Measuring Reward Hacking in Long-Horizon Coding Agents

DGX agent

arXiv:2605.21384v1 Announce Type: cross Abstract: As long-horizon coding agents produce more code than any developer can review, oversight collapses onto a single surface: the automated test suite. Re

model-releasesarxiv-cs-cl
21 May 2026
Safety

Stage-Audit: Auditable Source-Frontier Discovery for Cross-Wiki Tables

DGX agent

arXiv:2605.20478v1 Announce Type: new Abstract: LLM-curated tables can appear source-grounded while containing unsupported rows: the curator may recall entries from parametric memory and retroactively

safetyarxiv-cs-cl
21 May 2026
Research

Strategy-Induct: Task-Level Strategy Induction for Instruction Generation

DGX agent

arXiv:2605.20924v1 Announce Type: new Abstract: Designing effective task-level prompts is crucial for improving the performance of Large Language Models (LLMs). While prior work on instruction inducti

researcharxiv-cs-cl
21 May 2026
Model Releases

SymbolicLight V1: Spike-Gated Dual-Path Language Modeling with High Activation Sparsity and Sub-Billion-Scale Pre-Training Evidence

DGX agent

arXiv:2605.21333v1 Announce Type: new Abstract: Natively trained spiking language models struggle to combine Transformer-like language quality, stable multi-domain pre-training, and high activation sp

model-releasesarxiv-cs-cl
21 May 2026
Safety

Synchronization and Turn-Taking in Full-Duplex Speech Dialogue Models

DGX agent

arXiv:2605.20356v1 Announce Type: new Abstract: Full-duplex spoken dialogue models (SDMs) can listen and speak simultaneously, enabling interaction dynamics closer to human conversation than turn-base

safetyarxiv-cs-cl
21 May 2026
Safety

Task-conditioned probing of instruction-tuned multimodal LLMs: Region-specific brain alignment patterns under naturalistic stimuli

DGX agent

arXiv:2506.08277v3 Announce Type: replace-cross Abstract: Recent voxel-wise multimodal brain encoding studies have shown that multimodal large language models (MLLMs) exhibit a higher degree of brain

safetyarxiv-cs-cl
21 May 2026
← Previous
1…8586878889…162
Next →