AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlog
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,759 results
Agents

Optimized Three-Dimensional Photovoltaic Structures with LLM guided Tree Search

DGX agent

arXiv:2605.16191v1 Announce Type: new Abstract: We present a case study for how AI coding systems can be used to generate novel scientific hypotheses. We combine a generic coding agent (Google's AntiG

agentsarxiv-cs-cl
18 May 2026
Model Releases
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

PerfCodeBench: Benchmarking LLMs for System-Level High-Performance Code Optimization

DGX agent

arXiv:2605.15222v1 Announce Type: cross Abstract: Large language models (LLMs) can often generate functionally correct code, but their ability to produce efficient implementations for performance-crit

model-releasesarxiv-cs-cl
18 May 2026
Research

Prompt Stability Scoring for Text Annotation with Large Language Models

DGX agent

arXiv:2407.02039v3 Announce Type: replace Abstract: Researchers are increasingly using language models (LMs) for text annotation. These approaches rely only on a prompt telling the model to return a g

researcharxiv-cs-cl
18 May 2026
Safety

PSD: Pushing the Pareto Frontier of Diffusion LLMs via Parallel Speculative Decoding

DGX agent

arXiv:2605.15609v1 Announce Type: new Abstract: Diffusion large language models (dLLMs) generate text by iteratively denoising masked token sequences. Although dLLMs can predict all masked positions i

safetyarxiv-cs-cl
18 May 2026
Model Releases

RapidUn: Influence-Driven Parameter Reweighting for Efficient Large Language Model Unlearning

DGX agent

arXiv:2512.04457v2 Announce Type: replace Abstract: Removing specific data influence from large language models (LLMs) remains challenging, as retraining is costly and existing approximate unlearning

model-releasesarxiv-cs-cl
18 May 2026
Research

Reasoning Models Don't Just Think Longer, They Move Differently

DGX agent

arXiv:2605.15454v1 Announce Type: new Abstract: Reasoning-trained language models often spend more tokens on harder problems, but longer chains of thought do not show whether a model is merely computi

researcharxiv-cs-cl
18 May 2026
Safety

Reference Games as a Testbed for the Alignment of Model Uncertainty and Clarification Requests

DGX agent

arXiv:2601.07820v2 Announce Type: replace Abstract: In human conversation, both interlocutors play an active role in maintaining mutual understanding. When listeners are uncertain about what speakers

safetyarxiv-cs-cl
18 May 2026
Safety

Response-Conditioned Parallel-to-Sequential Orchestration for Multi-Agent Systems

DGX agent

arXiv:2605.15573v1 Announce Type: new Abstract: Multi-agent systems can solve complex tasks through collaboration between multiple Large Language Model agents. Existing collaboration frameworks typica

safetyarxiv-cs-cl
18 May 2026
Model Releases

SGR: A Stepwise Reasoning Framework for LLMs with External Subgraph Generation

DGX agent

arXiv:2605.16117v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated strong capabilities across diverse NLP applications, such as translation, text generation, and question a

model-releasesarxiv-cs-cl
18 May 2026
Model Releases

SMMBench: A Benchmark for Source-Distributed Multimodal Agent Memory

DGX agent

arXiv:2605.15710v1 Announce Type: new Abstract: Existing benchmarks for multimodal memory reasoning largely evaluate systems within pre-assembled contexts, but under-evaluate whether agents can use ev

model-releasesarxiv-cs-cl
18 May 2026
Research

Smoothie: Smoothing Diffusion on Token Embeddings for Text Generation

DGX agent

arXiv:2505.18853v2 Announce Type: replace Abstract: Diffusion models have achieved state-of-the-art performance in generating images, audio, and video, but their adaptation to text remains challenging

researcharxiv-cs-cl
18 May 2026
Research

Stabilizing Knowledge, Promoting Reasoning: Dual-Token Constraints for RLVR

DGX agent

arXiv:2507.15778v2 Announce Type: replace Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has become an effective post-training method for improving the reasoning abilities of Large La

researcharxiv-cs-cl
18 May 2026
Model Releases

STS: Efficient Sparse Attention with Speculative Token Sparsity

DGX agent

arXiv:2605.15508v1 Announce Type: cross Abstract: The quadratic complexity of attention imposes severe memory and computational bottlenecks on Large Language Model (LLM) inference. This challenge is p

model-releasesarxiv-cs-cl
18 May 2026
Research

Syntax Without Semantics: Teaching Large Language Models to Code in an Unseen Language

DGX agent

arXiv:2605.15607v1 Announce Type: new Abstract: Large language models (LLMs) achieve high pass rates on code generation benchmarks, yet whether they can transfer this ability to languages absent from

researcharxiv-cs-cl
18 May 2026
Safety

TemplateRL: Structured Template-Guided Reinforcement Learning for LLM Reasoning

DGX agent

arXiv:2505.15692v5 Announce Type: replace Abstract: Reinforcement learning (RL) has emerged as an effective paradigm for enhancing model reasoning. However, existing RL methods like GRPO typically rel

safetyarxiv-cs-cl
18 May 2026
Research

Toward LLMs Beyond English-Centric Development

DGX agent

arXiv:2605.15613v1 Announce Type: new Abstract: Through an analysis of sequences generated by open-weight large language models (LLMs), we demonstrate that LLMs are heavily biased toward English. Whil

researcharxiv-cs-cl
18 May 2026
Model Releases

VCG-Bench: Towards A Unified Visual-Centric Benchmark for Structured Generation and Editing

DGX agent

arXiv:2605.15677v1 Announce Type: new Abstract: Despite the rapid advancements in Vision-Language Models (VLMs), a critical gap remains in their ability to handle structured, controllable diagrammatic

model-releasesarxiv-cs-cl
18 May 2026
Safety

VSPO: Vector-Steered Policy Optimization for Behavioral Control

DGX agent

arXiv:2605.15604v1 Announce Type: cross Abstract: Modern language models often need to optimize a primary accuracy objective while also accommodating secondary behavioral preferences, such as verbosit

safetyarxiv-cs-cl
18 May 2026
Safety

When Importance Sampling Misallocates Credit: Asymmetric Ratios for Outcome-Supervised RL

DGX agent

arXiv:2510.06062v2 Announce Type: replace Abstract: Reinforcement learning (RL) has shown great promise in large language models (LLMs) post-training, which typically rely on token-level clipping to m

safetyarxiv-cs-cl
18 May 2026
Safety

When Latent Geometry Is Not Enough: Draft-Conditioned Latent Refinement for Non-Autoregressive Text Generation

DGX agent

arXiv:2605.15557v1 Announce Type: new Abstract: Continuous diffusion and flow models are attractive for non-autoregressive text generation because they can update all positions in parallel. A major di

safetyarxiv-cs-cl
18 May 2026
Research

Why are language models less surprised than humans? Testing the Parse Multiplicity Mismatch Hypothesis

DGX agent

arXiv:2605.15440v1 Announce Type: new Abstract: Surprisal theory posits that the processing difficulty of a word is determined by its predictability in context, offering a potential link between human

researcharxiv-cs-cl
18 May 2026
Model Releases

A Calculus-Based Framework for Determining Vocabulary Size in End-to-End ASR

DGX agent

arXiv:2605.14427v1 Announce Type: new Abstract: In hybrid automatic speech recognition (ASR) systems, the vocabulary size is unambiguous, typically determined by the number of phones, bi-phones, or tr

model-releasesarxiv-cs-cl
15 May 2026
Research

A Formative Study of Brief Affective Text as a Complement to Wearable Sensing for Longitudinal Student Health Monitoring

DGX agent

arXiv:2605.14360v1 Announce Type: cross Abstract: Wearable devices capture physiological and behavioral data with increasing fidelity, but the psychological context shaping these outcomes is difficult

researcharxiv-cs-cl
15 May 2026
Research

A Hormone-inspired Emotion Layer for Transformer language models (HELT)

DGX agent

arXiv:2605.13858v1 Announce Type: cross Abstract: Large Language Models have demonstrated remarkable capabilities in generating contextually relevant and grammatically correct text. However, they fund

researcharxiv-cs-cl
15 May 2026
Model Releases

A Large Language Model Based Pipeline for Review of Systems Entity Recognition from Clinical Notes

DGX agent

arXiv:2506.11067v3 Announce Type: replace Abstract: Objective: Develop a cost-effective, large language model (LLM)-based pipeline for automatically extracting Review of Systems (ROS) entities from cl

model-releasesarxiv-cs-cl
15 May 2026
Model Releases

Auditing Agent Harness Safety

DGX agent

arXiv:2605.14271v1 Announce Type: new Abstract: LLM agents increasingly run inside execution harnesses that dispatch tools, allocate resources, and route messages between specialized components. Howev

model-releasesarxiv-cs-cl
15 May 2026
Model Releases

Automated Construction of a Knowledge Graph of Nuclear Fusion Energy for Effective Elicitation and Retrieval of Information

DGX agent

arXiv:2504.07738v3 Announce Type: replace Abstract: In this document, we discuss a multi-step approach to automated construction of a knowledge graph, for structuring and representing domain-specific

model-releasesarxiv-cs-cl
15 May 2026
Research

Beyond Cosine Similarity: Zero-Initialized Residual Complex Projection for Aspect-Based Sentiment Analysis

DGX agent

arXiv:2603.28205v2 Announce Type: replace Abstract: Aspect-Based Sentiment Analysis (ABSA) faces critical challenges due to representation entanglement and false-negative collisions in real-valued emb

researcharxiv-cs-cl
15 May 2026
Model Releases

Beyond Mode-Seeking RL: Trajectory-Balance Post-Training for Diffusion Language Models

DGX agent

arXiv:2605.13935v1 Announce Type: cross Abstract: Diffusion language models are a promising alternative to autoregressive models, yet post-training methods for them largely adapt reward-maximizing obj

model-releasesarxiv-cs-cl
15 May 2026
Research

BOOKMARKS: Efficient Active Storyline Memory for Role-playing

DGX agent

arXiv:2605.14169v1 Announce Type: new Abstract: Memory systems are critical for role-playing agents (RPAs) to maintain long-horizon consistency. However, existing RPA memory methods (e.g., profiling)

researcharxiv-cs-cl
15 May 2026
Model Releases

Chain-of-Procedure: Hierarchical Visual-Language Reasoning for Procedural QA

DGX agent

arXiv:2605.14928v1 Announce Type: new Abstract: Recent advances in vision-language models (VLMs) have achieved impressive results on standard image-text tasks, yet their potential for visual procedure

model-releasesarxiv-cs-cl
15 May 2026
Safety

Comparing Developer and LLM Biases in Code Evaluation

DGX agent

arXiv:2603.24586v2 Announce Type: replace-cross Abstract: As LLMs are increasingly used as judges in code applications, they should be evaluated in realistic interactive settings that capture partial

safetyarxiv-cs-cl
15 May 2026
Research

Confidence Estimation for LLMs in Multi-turn Interactions

DGX agent

arXiv:2601.02179v2 Announce Type: replace Abstract: While confidence estimation is a promising direction for mitigating hallucinations in Large Language Models (LLMs), current research overwhelmingly

researcharxiv-cs-cl
15 May 2026
Research

Conversion of Lexicon-Grammar tables to LMF. Application to French

DGX agent

arXiv:2605.14816v1 Announce Type: new Abstract: We describe the first experiment of conversion of Lexicon-Grammar tables for French verbs into the Lexical Markup Framework (LMF) format. The Lexicon-Gr

researcharxiv-cs-cl
15 May 2026
Model Releases

CounselBench: A Large-Scale Expert Evaluation and Adversarial Benchmarking of Large Language Models in Mental Health Question Answering

DGX agent

arXiv:2506.08584v4 Announce Type: replace Abstract: Medical question answering (QA) benchmarks often focus on multiple-choice or fact-based tasks, leaving open-ended answers to real patient questions

model-releasesarxiv-cs-cl
15 May 2026
Research

Cross-Linguistic Transcription and Phonological Representation in the Huitongguanxi Huayiyiyu

DGX agent

arXiv:2605.14480v1 Announce Type: new Abstract: Purpose: This study investigates the transcription principles underlying Huitongguanxi Huayiyiyu (HHY), a series of multilingual glossaries compiled by

researcharxiv-cs-cl
15 May 2026
Safety

Distribution Corrected Offline Data Distillation for Large Language Models

DGX agent

arXiv:2605.14071v1 Announce Type: new Abstract: Distilling reasoning traces from strong large language models into smaller ones is a promising route to improve intelligence in resource-constrained set

safetyarxiv-cs-cl
15 May 2026
Research

Do Composed Image Retrieval Benchmarks Require Multimodal Composition?

DGX agent

arXiv:2605.14787v1 Announce Type: cross Abstract: Composed Image Retrieval (CIR) is a multimodal retrieval task where a query consists of a reference image and a textual modification, and the goal is

researcharxiv-cs-cl
15 May 2026
Safety

Do Reasoning LLMs Refuse What They Infer in Long Contexts?

DGX agent

arXiv:2602.08874v2 Announce Type: replace Abstract: Long-context LLMs can infer objectives that are not stated explicitly. This capability is useful for reasoning over documents, code, retrieved evide

safetyarxiv-cs-cl
15 May 2026
Research

Does Local News Stay Local?: Online Content Shifts in Sinclair-Acquired Stations

DGX agent

arXiv:2510.07060v2 Announce Type: replace Abstract: Local news stations are often considered to be reliable sources of non-politicized information, particularly local concerns that residents care abou

researcharxiv-cs-cl
15 May 2026
Applications

DT-Transformer: A Foundation Model for Disease Trajectory Prediction on a Real-world Health System

DGX agent

arXiv:2605.14227v1 Announce Type: cross Abstract: Accurate disease trajectory prediction is critical for early intervention, resource allocation, and improving long-term outcomes. While electronic hea

applicationsarxiv-cs-cl
15 May 2026
Safety

Dual Hierarchical Dialogue Policy Learning for Legal Inquisitive Conversational Agents

DGX agent

arXiv:2605.14057v1 Announce Type: new Abstract: Most existing dialogue systems are user-driven, primarily designed to fulfill user requests. However, in many critical real-world scenarios, a conversat

safetyarxiv-cs-cl
15 May 2026
Model Releases

EndPrompt: Efficient Long-Context Extension via Terminal Anchoring

DGX agent

arXiv:2605.14589v1 Announce Type: new Abstract: Extending the context window of large language models typically requires training on sequences at the target length, incurring quadratic memory and comp

model-releasesarxiv-cs-cl
15 May 2026
Applications

Energy-Regularized Sequential Model Editing on Hyperspheres

DGX agent

arXiv:2510.01172v3 Announce Type: replace Abstract: Large language models (LLMs) require constant updates to remain aligned with evolving real-world knowledge. Model editing offers a lightweight alter

applicationsarxiv-cs-cl
15 May 2026
Research

FactNet: A Billion-Scale Knowledge Graph for Multilingual Factual Grounding

DGX agent

arXiv:2602.03417v2 Announce Type: replace Abstract: Large language models hallucinate factual claims and struggle to ground their outputs in retrievable evidence, particularly in non-English languages

researcharxiv-cs-cl
15 May 2026
Research

Factorization-Error-Free Discrete Diffusion Language Model via Speculative Decoding

DGX agent

arXiv:2605.14305v1 Announce Type: new Abstract: Discrete diffusion language models improve generation efficiency through parallel token prediction, but standard X_0 prediction methods introduce factor

researcharxiv-cs-cl
15 May 2026
Model Releases

Forgetting That Sticks: Quantization-Permanent Unlearning via Circuit Attribution

DGX agent

arXiv:2605.15138v1 Announce Type: cross Abstract: Standard unlearning evaluations measure behavioral suppression in full precision, immediately after training, despite every deployed language model be

model-releasesarxiv-cs-cl
15 May 2026
Model Releases

From Scenes to Elements: Multi-Granularity Evidence Retrieval for Verifiable Multimodal RAG

DGX agent

arXiv:2605.15019v1 Announce Type: new Abstract: Multimodal Retrieval-Augmented Generation (RAG) systems retrieve evidence at coarse granularities (entire images or scenes), creating a mismatch with fi

model-releasesarxiv-cs-cl
15 May 2026
← Previous
1…9293949596…162
Next →