AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
4 Jun 2026

Cross-Prompt Generalization in Detecting AI-Generated Fake News Using Interpretable Linguistic Features

ResearchDGX agent

arXiv:2606.04199v1 Announce Type: new Abstract: The increasing use of large language models has raised concerns about the spread of AI-generated fake news, particularly under varying prompting strateg

CYGNET: Cypher Gate for Neural Execution Triage and Cost Containment

ApplicationsDGX agent

arXiv:2606.04645v1 Announce Type: new Abstract: Language models acting as agents over knowledge graphs generate Cypher queries that fail structurally (crashing at the database) or semantically (execut

Data Attribution in Large Language Models via Bidirectional Gradient Optimization

ResearchDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2606.04928v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly deployed across diverse applications, raising critical questions for governance, accountability, and dat

DEER: Disentangled Mixture of Experts with Instance-Adaptive Routing for Generalizable Machine-Generated Text Detection

ResearchDGX agent

arXiv:2511.01192v2 Announce Type: replace Abstract: Detecting machine-generated text has become a critical challenge amid the rapid advancement of LLMs, yet existing detectors degrade severely under d

Deliberate Evolution: Agentic Reasoning for Sample-Efficient Symbolic Regression with LLMs

AgentsDGX agent

arXiv:2606.04360v1 Announce Type: new Abstract: Symbolic regression (SR) discovers compact mathematical expressions from data, yet recent LLM-based evolutionary methods remain sample-inefficient becau

Depth-Attention: Cross-Layer Value Mixing for Language Models

ResearchDGX agent

arXiv:2606.05014v1 Announce Type: new Abstract: Self-attention selects information freely across the sequence, but across depth, Transformers merely add each layer's output to the residual stream, so

Discourse-Role Labels as Presentation-Time Variables for Context Use in Language Models

Model ReleasesDGX agent

arXiv:2606.04109v1 Announce Type: new Abstract: Context-augmented language model systems often wrap supplied content with labels such as Reference:, Evidence:, Instruction:, Note:, or Example:, but th

Disentangling Answer Engine Optimization from Platform Growth: A Log-Based Natural Experiment on ChatGPT Referral Traffic

ResearchDGX agent

arXiv:2606.04362v1 Announce Type: cross Abstract: Large language model (LLM) 'answer engines' such as ChatGPT now send measurable referral traffic to the open web, and a practice analogous to search e

DLLG: Dynamic Logit-Level Gating of LLM Experts

Model ReleasesDGX agent

arXiv:2606.04378v1 Announce Type: new Abstract: Leveraging multiple specialized LLMs can combine complementary strengths, but existing approaches trade adaptability for stability: routing commits prem

DuDi: Dual-Signal Distillation with Cross-Lingual Verbalizer

SafetyDGX agent

arXiv:2606.04694v1 Announce Type: new Abstract: Small language models (SLMs) are efficient and scalable, but their multilingual capabilities degrade severely at sub-billion scales, especially for Sout

Effective vocabulary expansion of multilingual language models for extremely low-resource languages

TutorialsDGX agent

arXiv:2602.09388v2 Announce Type: replace Abstract: Multilingual pre-trained language models(mPLMs) offer significant benefits for many low-resource languages. To further expand the range of languages

Efficient Reasoning on the Edge

Local AiDGX agent

arXiv:2603.16867v2 Announce Type: replace-cross Abstract: Large language models (LLMs) with chain-of-thought reasoning achieve state-of-the-art performance across complex problem-solving tasks, but th

Enhancing Hallucination Detection through Noise Injection

ResearchDGX agent

arXiv:2502.03799v4 Announce Type: replace Abstract: Large Language Models (LLMs) are prone to generating plausible yet incorrect responses, known as hallucinations. Effectively detecting hallucination

Entity Binding Failures in Speech LLM Reasoning: Diagnosis and Chain-of-Thought Intervention

ResearchDGX agent

arXiv:2606.04474v1 Announce Type: new Abstract: Speech Large Language Models (SLLMs) underperform their text counterparts on complex reasoning. We reveal that this modality gap is not a uniform cognit

Evaluating Autoformalization Robustness via Semantically Similar Paraphrasing

ResearchDGX agent

arXiv:2511.12784v3 Announce Type: replace Abstract: Large Language Models (LLMs) have recently emerged as powerful tools for autoformalization. Despite their impressive performance, these models can s

Evaluating Large Language Models in Dynamic Clinical Decision-Making with Standardized Patient Cases

Model ReleasesDGX agent

arXiv:2606.05112v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly proposed as clinical agents, yet static, single-turn benchmarks cannot capture how a model dynamically del

Expert-Aware Refusal Steering

SafetyDGX agent

arXiv:2606.04160v1 Announce Type: new Abstract: Safety alignment in instruction-tuned large language models (LLMs) depends on a model's ability to reliably refuse to respond to harmful or disallowed r

Exploring the Topology and Memory of Consensus: How LLM Agents Agree, Fragment, or Settle When Forming Conventions

Model ReleasesDGX agent

arXiv:2606.04197v1 Announce Type: cross Abstract: How much should an LLM agent remember, and how should multi-agent systems be connected when trying to reach consensus? We show these two design choice

Fast & Faithful Function Vectors

ResearchDGX agent

arXiv:2606.05079v1 Announce Type: new Abstract: Function vectors (FVs) are task representations elicited during in-context learning that can be used to steer Large Language Models (LLMs). However, des

Few Tokens, Big Leverage: Preserving Safety Alignment by Constraining Safety Tokens during Fine-tuning

SafetyDGX agent

arXiv:2603.07445v2 Announce Type: replace Abstract: Large language models (LLMs) often require fine-tuning (FT) to perform well on downstream tasks, but FT can induce safety-alignment drift even when

Fine-grained Fragment Retrieval in Multi-modal Long-form Dialogues

ApplicationsDGX agent

arXiv:2606.04591v1 Announce Type: new Abstract: With the widespread adoption of multi-modal communication platforms, long-form dialogues interleaving text and images have become increasingly common. U

GARL: Game-Theoretic Reinforcement Learning for Multi-Agent Strategic Prioritisation

SafetyDGX agent

arXiv:2606.05002v1 Announce Type: new Abstract: LLM-based multi-agent systems are increasingly used for strategic decision-making tasks. In such settings, performance depends not only on individual mo

GENEB: Why Genomic Models Are Hard to Compare

Model ReleasesDGX agent

arXiv:2606.04525v1 Announce Type: new Abstract: Progress in genomic foundation models is difficult to assess due to fragmented benchmarks, incompatible evaluation protocols, and task-specific reportin

GIFT: Games as Informal Training for Generalizable LLMs

ResearchDGX agent

arXiv:2601.05633v2 Announce Type: replace Abstract: Recent LLMs excel at formal tasks such as mathematical reasoning and code generation, but still struggle with broader abilities such as planning, cr

Global Sketch-Based Watermarking for Diffusion Language Models

SafetyDGX agent

arXiv:2606.04486v1 Announce Type: cross Abstract: Watermarking methods for language models have been studied extensively in the autoregressive setting, where tokens are generated sequentially. These w

GlossAssist -- A Tool to Simplify Corpus Creation and Study the Effect of NLP Models in Low-Resource Documentation Settings

ResearchDGX agent

arXiv:2606.04367v1 Announce Type: new Abstract: Interlinear glossed text (IGT) is the standard format for linguistic annotation in language documentation. Producing it manually, however, is often slow

GRAIL: Gradient-Reweighted Advantages for Reinforcement Learning with Verifiable Rewards

SafetyDGX agent

arXiv:2606.04889v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (e.g. GRPO) is now a common way to improve mathematical reasoning in Large Language Models (LLMs). Howeve

Graph-R1: Towards Agentic GraphRAG Framework via End-to-end Reinforcement Learning

AgentsDGX agent

arXiv:2507.21892v2 Announce Type: replace Abstract: Retrieval-Augmented Generation (RAG) mitigates hallucination in LLMs by incorporating external knowledge, but relies on chunk-based retrieval that l

High-Quality Entity Segmentation and Grounding

Model ReleasesDGX agent

arXiv:2402.02555v2 Announce Type: replace-cross Abstract: In this work, we propose ESG, a pipeline for high-quality entity segmentation and grounding supported by a new dataset EntitySeg. At first, th

Hybrid Adversarial Defence for Natural Language Understanding Tasks

SafetyDGX agent

arXiv:2606.04612v1 Announce Type: new Abstract: Large Language Models (LLMs) are vulnerable both to hallucination and adversarial manipulation. Although these problems are closely related, existing de

Imbuing Large Language Models with Bidirectional Logic for Robust Chain Repair

SafetyDGX agent

arXiv:2606.05030v1 Announce Type: new Abstract: Autoregressive chain-of-thought (CoT) reasoning in large language models (LLMs) is fundamentally forward-directed: each step conditions only on prior to

In-Context Graphical Inference

SafetyDGX agent

arXiv:2606.05042v1 Announce Type: cross Abstract: Marginal inference in discrete graphical models forces a choice between exactness and scalability: exact algorithms are intractable for high-treewidth

Large Language Models in K-12 Education: Alignment with State Curriculum Standards and Student Personas

SafetyDGX agent

arXiv:2606.04846v1 Announce Type: new Abstract: As Large Language Models (LLMs) become increasingly popular in educational settings, they raise important questions about the ethical implications of th

LazyAttention: Efficient Retrieval-Augmented Generation with Deferred Positional Encoding

ResearchDGX agent

arXiv:2606.04302v1 Announce Type: new Abstract: Key-value (KV) caching accelerates inference of large language models (LLMs) by reusing past computations for generated tokens. Its importance becomes e

LDARNet: DNA Adaptive Representation Network with Learnable Tokenization for Genomic Modeling

Model ReleasesDGX agent

arXiv:2606.04552v1 Announce Type: new Abstract: Genomic foundation models increasingly adopt large language model architectures, yet almost universally rely on fixed tokenization schemes such as k-mer

Learning What to Learn: Stage-Specific Data Sets for SFT-then-RL in Small Language Model Reasoning

TutorialsDGX agent

arXiv:2606.04466v1 Announce Type: new Abstract: Post-training Small Language Models (SLMs) for reasoning typically follows an SFT-then-RL pipeline, yet existing work rarely considers what data should

Learning When to Act or Refuse: Guarding Agentic Reasoning Models for Safe Multi-Step Tool Use

Model ReleasesDGX agent

arXiv:2603.03205v2 Announce Type: replace Abstract: Agentic language models operate in a fundamentally different safety regime than chat models: they must plan, call tools, and execute long-horizon ac

LifeSide: Benchmarking Agents as Lifelong Digital Companions

Model ReleasesDGX agent

arXiv:2606.04660v1 Announce Type: new Abstract: Lifelong digital companions must integrate cross-session cues, continually update their understanding of users, and adapt to shifting privacy boundaries

Light or Full Verb? A Minimal-Pair Dataset for Probing Phraseological Competence in Language Models

ResearchDGX agent

arXiv:2606.05087v1 Announce Type: new Abstract: Frequent English verbs such as 'have' and 'make' can function either as collocates in light-verb constructions or as full lexical predicates, as in 'mak

LiSeCo: Linear Semantic Control for Language Generation

ResearchDGX agent

arXiv:2405.15454v4 Announce Type: replace Abstract: The prevalence of Large Language Models (LLMs) in critical applications highlights the need for controlled language generation methods that are both

Listening to the Workforce: Measuring Construction Worker Safety Attitudes from Social Media Discourse Using LLMs

SafetyDGX agent

arXiv:2606.04450v1 Announce Type: new Abstract: Worker safety attitudes are key determinants of whether protective practices are applied or bypassed on construction sites. Yet measuring them at scale

LLMs + Persona-Plug = Personalized LLMs

Model ReleasesDGX agent

arXiv:2409.11901v2 Announce Type: replace Abstract: Personalization plays a critical role in numerous language tasks and applications, since users with the same requirements may prefer diverse outputs

Long Live Fine-Tuning: Task-Specific Transformers Outperform Zero-Shot LLMs for Misinformation Response Classification on Reddit

Model ReleasesDGX agent

arXiv:2606.04274v1 Announce Type: new Abstract: As large language models (LLMs) become default tools for online information verification, an implicit assumption follows them: that scale and general ca

Multilingual Long-Form Speech Instruction Following: KIT's Submission to IWSLT 2026

ResearchDGX agent

arXiv:2606.04730v1 Announce Type: new Abstract: With the advent of Large Language Models, single-task and token-based multi-task models have evolved into instruction-based systems that infer task and

MusaCoder: Native GPU Kernel Generation with Full-Stack Training on Moore Threads GPU

SafetyDGX agent

arXiv:2606.04847v1 Announce Type: cross Abstract: Native GPU kernel generation turns high-level tensor programs into executable, efficient low-level code. Existing Large Language Models (LLMs) struggl

NextMotionQA: Benchmarking and Judging Human Motion Understanding with Vision-Language Models

Model ReleasesDGX agent

arXiv:2606.04773v1 Announce Type: cross Abstract: Reliable evaluation of human motion understanding is fundamental to advancing embodied AI, robotics, and animation. However, existing benchmarks suffe

Noisy memory encoding explains negative polarity illusions

ResearchDGX agent

arXiv:2606.04340v1 Announce Type: new Abstract: A sentence like 'The authors that no critics recommended have ever received acknowledgment for a best-selling novel' is sometimes rated as acceptable ev

Off-Distribution Voices: Fanfiction Subgenres as Universal Vernacular Jailbreaks for Aligned LLMs

SafetyDGX agent

arXiv:2606.04483v1 Announce Type: new Abstract: Existing jailbreaks against aligned LLMs are discrete artifacts whose surface forms are easy to fingerprint and patch. We argue that the real failure mo

Optimizing the Cost-Quality Tradeoff of Agentic Theorem Provers in Lean

AgentsDGX agent

arXiv:2606.04883v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used in workflows for generating formal proofs in Lean. These workflows often decompose problems into smal

Outcome-Grounded Advantage Reshaping for Fine-Grained Credit Assignment in Mathematical Reasoning

SafetyDGX agent

arXiv:2601.07408v2 Announce Type: replace Abstract: Group Relative Policy Optimization (GRPO) has emerged as a promising critic-free reinforcement learning paradigm for reasoning tasks. However, stand

Parameter-Efficient Fine-Tuning with Learnable Rank

Model ReleasesDGX agent

arXiv:2606.04325v1 Announce Type: new Abstract: Low-Rank Adaptation (LoRA) is a popular parameter-efficient fine-tuning (PEFT) method that restricts weight updates to low-rank adapters, introducing a

PersonaTree: Structured Lifecycle Memory for Person Understanding in LLM Agents

SafetyDGX agent

arXiv:2606.04780v1 Announce Type: new Abstract: Persistent LLM agents require memory representations that make the formation of person understanding explicit across long term interaction. Existing age

Physics-Informed Neural Network Modeling of Biodegradable Contaminant Transport through GCL/SL Composite Liners

ResearchDGX agent

arXiv:2606.04392v1 Announce Type: cross Abstract: This study develops a two-domain physics-informed neural network framework for contaminant transport through a GCL/SL composite liner system, in which

Probing Outcome-Level Resemblance and Mechanism-Level Alignment in LLM Risk Decisions: Evidence from the St. Petersburg Game

SafetyDGX agent

arXiv:2606.04978v1 Announce Type: new Abstract: LLMs can appear cautious in risk decision-making tasks, yet cautious-looking outputs do not necessarily indicate alignment with human decision-making me

Prompt-Level Distillation: A Non-Parametric Alternative to Model Fine-Tuning for Efficient Reasoning

Model ReleasesDGX agent

arXiv:2602.21103v2 Announce Type: replace Abstract: Advanced reasoning typically requires Chain-of-Thought prompting, which is accurate but incurs prohibitive latency and substantial test-time inferen

Query-based Cross-Modal Projector Bolstering Mamba Multimodal LLM

ResearchDGX agent

arXiv:2606.04719v1 Announce Type: new Abstract: The Transformer's quadratic complexity with input length imposes an unsustainable computational load on large language models (LLMs). In contrast, the S

RAMPART: Registry-based Agentic Memory with Priority-Aware Runtime Transformation

Model ReleasesDGX agent

arXiv:2606.04628v1 Announce Type: new Abstract: RAMPART is a compile-time memory model and pure in-RAM block registry for LLM-based agents. Context assembly is a programmable runtime operation where c

Read the Trace, Steer the Path: Trajectory-Aware Reinforcement Learning for Diffusion Language Models

ResearchDGX agent

arXiv:2606.04396v1 Announce Type: new Abstract: Diffusion large language models (dLLMs) generate responses by iteratively unmasking and revising many positions in parallel. This process leaves a rich

Read What You Hear: Reference-Free Hypotheses Evaluation with Acoustic Discrepancy

ResearchDGX agent

arXiv:2606.04680v1 Announce Type: cross Abstract: Automatic speech recognition systems commonly rely on reference transcriptions for evaluation, while reference-free approaches often depend on interna

Reasoning over Boundaries: Enhancing Specification Alignment via Test-time Deliberation

Model ReleasesDGX agent

arXiv:2509.14760v3 Announce Type: replace Abstract: Large language models (LLMs) are increasingly applied in diverse real-world scenarios, each governed by bespoke behavioral and safety specifications

← Previous
1…4445464748…129
Next →