AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlog
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,759 results
Research

When RL Suppresses Its Own Vocabulary: Recovering Reasoning Diversity in Puzzle-to-Math Transfer

DGX agent

arXiv:2605.29190v1 Announce Type: cross Abstract: Reinforcement learning using verifiable rewards (RLVR) improves LLM reasoning, but the conditions under which it transfers across domains -- and why i

researcharxiv-cs-cl
29 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

When the Same Coefficients Reach Different Places: Asymmetric Realizability in Transplanting Tokenizers across Large Language Models

DGX agent

arXiv:2601.00065v3 Announce Type: replace-cross Abstract: Tokenizer transplant in cross-vocabulary model composition reconstructs donor-only embedding rows as weighted combinations over shared lexical

model-releasesarxiv-cs-cl
29 May 2026
Applications

Who Am I? History-Aware Profiles for Student Simulation in Tutoring Dialogues

DGX agent

arXiv:2605.30051v1 Announce Type: new Abstract: A key part of developing large language model (LLM)-powered, automated tutoring tools is student simulation, i.e., using LLMs to role-play as students,

applicationsarxiv-cs-cl
29 May 2026
Model Releases

World Models in Words: Auditing Physical State-Transition Commitments in Vision-Language Models

DGX agent

arXiv:2605.29585v1 Announce Type: new Abstract: Vision-language models (VLMs) are increasingly used to answer questions about physical scenes, yet most evaluations reduce performance to a final answer

model-releasesarxiv-cs-cl
29 May 2026
Agents

WorldMemArena: Evaluating Multimodal Agent Memory Through Action-World Interaction

DGX agent

arXiv:2605.29341v1 Announce Type: cross Abstract: Multimodal large language models are increasingly deployed as long-horizon agents, where memory must do more than recall: it must track an evolving wo

agentsarxiv-cs-cl
29 May 2026
Research

X-GS: An Extensible Framework for Perceiving and Thinking via 3D Gaussian Splatting

DGX agent

arXiv:2603.09632v3 Announce Type: replace-cross Abstract: 3D Gaussian Splatting (3DGS) has emerged as a powerful technique for novel view synthesis, subsequently extending into numerous spatial AI app

researcharxiv-cs-cl
29 May 2026
Research

A new semantically annotated corpus with syntactic-semantic and cross-lingual senses

DGX agent

arXiv:2605.28494v1 Announce Type: new Abstract: We describe a new sense-tagged corpus for word sense disambiguation. The corpus is constituted of instances of 20 French polysemous verbs. Each verb ins

researcharxiv-cs-cl
28 May 2026
Research

A tree interpretation of arc standard dependency derivation

DGX agent

arXiv:2603.27459v2 Announce Type: replace Abstract: Arc-standard derivations over projective dependency trees can be interpreted as the incremental construction of lexicalized ordered trees with conti

researcharxiv-cs-cl
28 May 2026
Applications

A Wolf in Sheep's Clothing: Targeted Routing Hijacking in Federated RAG

DGX agent

arXiv:2605.28112v1 Announce Type: cross Abstract: Federated Retrieval-Augmented Generation (FedRAG) is attractive for privacy-sensitive applications because raw data remain local. As a result, routing

applicationsarxiv-cs-cl
28 May 2026
Safety

Activation Steering for Synthetic Data Generation: The Role of Diversity in Downstream Safety Detection

DGX agent

arXiv:2605.28664v1 Announce Type: cross Abstract: Safety detection models require examples of HHH (Helpful, Harmless, Honest)-violating outputs for robust generalization, however such examples are sca

safetyarxiv-cs-cl
28 May 2026
Model Releases

AdaDPO: Self-Adaptive Direct Preference Optimization with Balanced Gradient Updates

DGX agent

arXiv:2605.28440v1 Announce Type: new Abstract: DPO has become a widely adopted alternative to RLHF for aligning LLMs with human preferences, eliminating the need for a separate reward model or RL loo

model-releasesarxiv-cs-cl
28 May 2026
Model Releases

Adaptive Cost-Efficient Evaluation for Reliable Patent Claim Generation

DGX agent

arXiv:2604.04295v3 Announce Type: replace Abstract: Automated patent claim validation demands low error tolerance. However, existing approaches face a rigidity-resource dilemma: lightweight encoders c

model-releasesarxiv-cs-cl
28 May 2026
Applications

Addressing Pitfalls in Auditing Practices of Automatic Speech Recognition Technologies: A Case Study of People with Aphasia

DGX agent

arXiv:2506.08846v3 Announce Type: replace-cross Abstract: Automatic Speech Recognition (ASR) systems' growing use warrants robust auditing approaches to ensure equitable transcription quality, especia

applicationsarxiv-cs-cl
28 May 2026
Model Releases

AdvJudge-Zero: Binary Decision Flips in LLM-as-a-Judge via Adversarial Control Tokens

DGX agent

arXiv:2512.17375v2 Announce Type: replace-cross Abstract: LLM-as-a-Judge systems supply the reward signal in modern RLHF and RLVR pipelines, but their binary verdict reduces to a single linear readout

model-releasesarxiv-cs-cl
28 May 2026
Safety

Agent Explorative Policy Optimization for Multimodal Agentic Reasoning

DGX agent

arXiv:2605.28774v1 Announce Type: new Abstract: Vision-language models with extended reasoning succeed on complex problems, but many real-world problems require external tools that internal reasoning

safetyarxiv-cs-cl
28 May 2026
Model Releases

Agentic Separation Logic Specification Synthesis

DGX agent

arXiv:2605.27531v1 Announce Type: cross Abstract: Specification synthesis, the task of automatically inferring formal specifications from program implementations and natural language, is important for

model-releasesarxiv-cs-cl
28 May 2026
Safety

Agents that Matter: Optimizing Multi-Agent LLMs via Removal-Based Attribution

DGX agent

arXiv:2605.27621v1 Announce Type: cross Abstract: As multi-agent systems (MAS) become increasingly complex, identifying the contributions of individual agents is critical for system optimization. Howe

safetyarxiv-cs-cl
28 May 2026
Agents

AI Research Agents Narrow Scientific Exploration

DGX agent

arXiv:2605.27905v1 Announce Type: new Abstract: AI research agents can now generate research ideas, design experiments, run code, and draft papers, raising the possibility of large-scale AI-assisted s

agentsarxiv-cs-cl
28 May 2026
Local Ai

An Evolutionary Approach for Designing Stable and Highly Expressible Low-Immunogenicity Therapeutic mRNA Sequences

DGX agent

arXiv:2605.27986v1 Announce Type: new Abstract: Messenger RNA (mRNA) sequences as therapeutics require optimized design to ensure efficient translation, structural stability, and minimal immunogenicit

local-aiarxiv-cs-cl
28 May 2026
Applications

Analyzing Cancer Patients' Experiences with Embedding-based Topic Modeling and LLMs

DGX agent

arXiv:2601.12154v2 Announce Type: replace Abstract: This study investigates the use of neural topic modeling and LLMs to uncover meaningful themes from patient storytelling data, to offer insights tha

applicationsarxiv-cs-cl
28 May 2026
Model Releases

Analyzing Quality-Latency-Resource Trade-offs in a Technical Documentation RAG Assistant Using LoRA Adaptation

DGX agent

arXiv:2605.28222v1 Announce Type: new Abstract: We study quality-latency-resource trade-offs in a documentation-grounded retrieval-augmented generation (RAG) system that uses Low-Rank Adaptation (LoRA

model-releasesarxiv-cs-cl
28 May 2026
Research

Are We Truly Innovating? A Qualitative and Quantitative Study of Originality in AI Research Papers

DGX agent

arXiv:2602.06054v3 Announce Type: replace Abstract: Assessing originality in AI research is arguably the most consequential yet least reliable step in peer review. Reviewer judgments of originality re

researcharxiv-cs-cl
28 May 2026
Model Releases

Argument Quality Assessment with Large Language Models: A Pairwise Bradley-Terry Approach

DGX agent

arXiv:2605.28313v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated remarkable capabilities in tasks related to reasoning and judgment. However, assessing the quality of arg

model-releasesarxiv-cs-cl
28 May 2026
Model Releases

Ask Now, Use Later: Benchmarking the Proactivity Gap in Long-Lived LLM Agents

DGX agent

arXiv:2605.28108v1 Announce Type: new Abstract: A long-lived LLM agent, such as OpenClaw, earns its value by acting on a user's preferences and constraints across sessions, not just the current reques

model-releasesarxiv-cs-cl
28 May 2026
Model Releases

Assessing Factual Music Comprehension in Large Audio Language Models

DGX agent

arXiv:2511.05550v2 Announce Type: replace-cross Abstract: Large audio language models (LALMs) leverage multimodal representations to generate open-ended answers to natural language queries about audio

model-releasesarxiv-cs-cl
28 May 2026
Model Releases

ATLAS: All-round Testing of Long-context Abilities across Scales

DGX agent

arXiv:2605.28079v1 Announce Type: new Abstract: Long-context language models now advertise context windows up to millions of tokens, yet evaluations typically report a single length or a narrow task f

model-releasesarxiv-cs-cl
28 May 2026
Research

Attention Projection Mixing with Exogenous Anchors

DGX agent

arXiv:2601.08131v4 Announce Type: replace Abstract: Cross-layer reuse of early attention projections can improve optimization and data efficiency, but it creates a structural conflict: the first layer

researcharxiv-cs-cl
28 May 2026
Safety

Auditing Stance Asymmetry in Generative Explanations

DGX agent

arXiv:2605.27988v1 Announce Type: new Abstract: Bias evaluation for language models has made substantial progress on bounded comparisons, such as overt derogation, stereotype association, or label-sen

safetyarxiv-cs-cl
28 May 2026
Applications

BEAR: Budgeted Evidence Allocation for Multi-Document Reasoning

DGX agent

arXiv:2601.18116v2 Announce Type: replace Abstract: We argue that multi-document reasoning is constrained not only by how much text a model can read, but also by how limited query-time evidence budget

applicationsarxiv-cs-cl
28 May 2026
Model Releases

Benchmarking and Mechanistic Analysis of Vision-Language Models for Cross-Depiction Assembly Instruction Alignment

DGX agent

arXiv:2604.00913v2 Announce Type: replace-cross Abstract: 2D assembly diagrams are often abstract and hard to follow, creating a need for intelligent assistants that can monitor progress, detect error

model-releasesarxiv-cs-cl
28 May 2026
Research

Better heads do not guarantee better binarized constituency parsing

DGX agent

arXiv:2605.28131v1 Announce Type: new Abstract: We revisit punctuation-aware tree binarization for constituency parsing and ask whether dependency-induced headedness improves binary parser supervision

researcharxiv-cs-cl
28 May 2026
Research

Beyond Chunk-Local Extraction: Cross-Chunk Graph Augmentation for GraphRAG

DGX agent

arXiv:2605.28004v1 Announce Type: new Abstract: GraphRAG extends retrieval-augmented generation by organizing corpora as explicit knowledge graphs, enabling graph-based retrieval for complex question

researcharxiv-cs-cl
28 May 2026
Research

Beyond Input Understanding: Diagnosing Multilingual Mathematical Reasoning with Directed Acyclic Trace Graphs

DGX agent

arXiv:2605.27715v1 Announce Type: new Abstract: Large reasoning models (LRMs) achieve strong mathematical reasoning performance in English, but remain much less reliable in many low- and medium-resour

researcharxiv-cs-cl
28 May 2026
Model Releases

Beyond One Path: Evaluating and Enhancing Divergent Thinking in Interactive LLM Agents

DGX agent

arXiv:2605.28465v1 Announce Type: new Abstract: Divergent thinking is a core dimension of creativity, yet existing evaluations of Large Language Models (LLMs) treat them as single-turn text generation

model-releasesarxiv-cs-cl
28 May 2026
Research

Beyond pass@k: Redundancy-Aware RLVR for Multi-Sample Code Generation

DGX agent

arXiv:2605.28022v1 Announce Type: new Abstract: LLMs for code generation are commonly evaluated in repeated-sampling settings using Pass@k, where multiple candidate programs are executed against unit

researcharxiv-cs-cl
28 May 2026
Safety

Boundary Suppression Asymmetry in Post-trained Assistants: Over-expansion as a Controllability Cost

DGX agent

arXiv:2605.27969v1 Announce Type: new Abstract: Post-trained language-model assistants are often optimized to avoid under-answering, encouraging complete, helpful, cautious, and proactive responses. W

safetyarxiv-cs-cl
28 May 2026
Safety

Breaking the Script Barrier: Enabling Automatic Alignment for PoS-based ASR Error Analysis in Non-Latin Scripts

DGX agent

arXiv:2605.28438v1 Announce Type: new Abstract: Automatic Speech Recognition (ASR) systems are commonly evaluated using aggregate metrics such as Word Error Rate (WER), which do not capture the lingui

safetyarxiv-cs-cl
28 May 2026
Model Releases

Building Community-Centred NLP Resources for Puno Quechua

DGX agent

arXiv:2605.28253v1 Announce Type: new Abstract: The preservation of under-resourced languages requires digital tools and resources shaped by and for their speakers. We present the first dedicated ASR

model-releasesarxiv-cs-cl
28 May 2026
Research

CALM-IT: Generating Realistic Long-Form Motivational Interviewing Dialogues with Dual-Actor Conversational Dynamics Tracking

DGX agent

arXiv:2601.10085v2 Announce Type: replace Abstract: Therapeutic dialogue is not a sequence of isolated responses: client goals, motivation, resistance, and therapeutic alliance evolve over time. Yet c

researcharxiv-cs-cl
28 May 2026
Model Releases

Camellia: Benchmarking Cultural Biases in LLMs for Asian Languages

DGX agent

arXiv:2510.05291v2 Announce Type: replace Abstract: As Large Language Models (LLMs) develop stronger multilingual capabilities, their sensitivity to culturally diverse entities becomes increasingly im

model-releasesarxiv-cs-cl
28 May 2026
Research

Can Hallucinations Be Useful? Solving Multi-Hop Questions With SLMs By Chaining System-I/II Reasoning

DGX agent

arXiv:2605.27596v1 Announce Type: new Abstract: Recently, there has been increased interest in Small Language Models (SLMs), which are fast, show good performance, and have lower hardware demands than

researcharxiv-cs-cl
28 May 2026
Model Releases

Can Large Language Models Handle Discourse Particles? A Case Study of Colloquial Malay

DGX agent

arXiv:2605.28782v1 Announce Type: new Abstract: Discourse particles, such as extit{well} and extit{kind of}, are crucial components that enable LLMs to ``speak'' more like humans. They are used to con

model-releasesarxiv-cs-cl
28 May 2026
Research

Can LLMs Use Linguistic Uncertainty Markers to Reliably Reflect Intrinsic Confidence?

DGX agent

arXiv:2605.28778v1 Announce Type: new Abstract: LLMs' linguistically expressed confidence should faithfully reflect their intrinsic uncertainty. While recent work shows LLMs struggle to use epistemic

researcharxiv-cs-cl
28 May 2026
Model Releases

CAREF: Calibration-Aware Regularization for Explanation Faithfulness Without Rationale Supervision

DGX agent

arXiv:2605.27835v1 Announce Type: cross Abstract: We introduce CAREF, a parameter-efficient fine-tuning framework that jointly optimizes predictive accuracy and explanation faithfulness via calibratio

model-releasesarxiv-cs-cl
28 May 2026
Agents

Chain-based Adaptive Reconfiguration Over Lattices for Hallucination Reduction

DGX agent

arXiv:2605.27706v1 Announce Type: new Abstract: We introduce CAROL (Chain-based Adaptive Reconfiguration Over Lattices), a probabilistic framework for test-time hallucination reduction in large langua

agentsarxiv-cs-cl
28 May 2026
Research

Challenges in Explaining Pretrained Clinical Text Classifiers

DGX agent

arXiv:2605.28060v1 Announce Type: new Abstract: Explaining the predictions of neural models in clinical NLP remains a significant challenge, especially for complex tasks involving long, unstructured m

researcharxiv-cs-cl
28 May 2026
Model Releases

Chinese Word Boundary Recovery through Character Alignment Projection

DGX agent

arXiv:2605.28128v1 Announce Type: new Abstract: Chinese word segmentation is especially fragile in non-standard text, where language learner errors and other character-level divergences disrupt the wo

model-releasesarxiv-cs-cl
28 May 2026
Safety

CIRF: Tokenizing Chain-of-Thoughts into Reusable Functional Units for Efficient Latent Reasoning in Large Language Models

DGX agent

arXiv:2605.28292v1 Announce Type: new Abstract: Implicit Chain-of-Thought (CoT) reduces the inference cost of large language models by internalizing the explicit rationales. However, existing approach

safetyarxiv-cs-cl
28 May 2026
← Previous
1…7071727374…162
Next →