AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
Model Releases

BreastGPT: A Multimodal Large Language Model for the Full Spectrum of Breast Cancer Clinical Routine

DGX agent

arXiv:2606.04911v1 Announce Type: cross Abstract: Breast cancer remains a leading cause of cancer-related mortality among women. Its clinical management requires multimodal reasoning across a clinical

model-releasesarxiv-cs-cl
4 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Caliper: Probing Lexical Anchors versus Causal Structure in LLMs

DGX agent

arXiv:2606.04915v1 Announce Type: new Abstract: Large language models reach 50 to 70% accuracy on causal reasoning benchmarks such as CLadder, but it is unclear whether this reflects structural reason

model-releasesarxiv-cs-cl
4 Jun 2026
Tutorials

Can Crowdsourcing Survive the LLM Era? A Community Survey on Human Data Collection

DGX agent

arXiv:2606.04924v1 Announce Type: new Abstract: The widespread use of Large Language Models (LLMs) as writing tools challenges the validity of crowdsourced data, as crowdworkers may outsource tasks to

tutorialsarxiv-cs-cl
4 Jun 2026
Model Releases

Can Large Language Models Generalize Procedures Across Representations?

DGX agent

arXiv:2602.03542v2 Announce Type: replace Abstract: Large language models (LLMs) are trained and tested extensively on symbolic representations such as code and graphs, yet real-world user tasks are o

model-releasesarxiv-cs-cl
4 Jun 2026
Hardware

Cartridges at Scale: Training Modular KV Caches over Large Document Collections

DGX agent

arXiv:2606.04557v1 Announce Type: new Abstract: Large Language Models can reason over long contexts, yet prefilling millions of tokens is wasteful as much of the content remains static across queries.

hardwarearxiv-cs-cl
4 Jun 2026
Tutorials

Characterizing, Evaluating, and Optimizing Complex Reasoning

DGX agent

arXiv:2602.08498v2 Announce Type: replace Abstract: Large Reasoning Models (LRMs) increasingly rely on reasoning traces with complex internal structures. However, existing work lacks a unified answer

tutorialsarxiv-cs-cl
4 Jun 2026
Research

CleanCodec: Efficient and Robust Speech Tokenization via Perceptually Guided Encoding

DGX agent

arXiv:2606.04418v1 Announce Type: cross Abstract: Neural audio codecs are a key component of speech processing pipelines, compressing audio into discrete tokens for downstream modeling. However, exist

researcharxiv-cs-cl
4 Jun 2026
Research

Clinical Assistant for Remote Engagement Link (CARE-link): A Web-Based Electronic Health Records Software for Managing Diabetes

DGX agent

arXiv:2606.04952v1 Announce Type: cross Abstract: CARE-link is an open-source, web-based clinical support platform designed to improve the management of gestational diabetes by linking clinicians and

researcharxiv-cs-cl
4 Jun 2026
Research

Computational conceptual history of scientific concepts: From early digital methods to LLMs

DGX agent

arXiv:2606.04118v1 Announce Type: new Abstract: This article situates large language models (LLMs) within the longer history of computational approaches to concept analysis in the history, philosophy,

researcharxiv-cs-cl
4 Jun 2026
Safety

Confidence Before Answering: A Paradigm Shift for Efficient LLM Uncertainty Estimation

DGX agent

arXiv:2603.05881v2 Announce Type: replace Abstract: Reliable deployment of large language models (LLMs) requires accurate uncertainty estimation. Existing methods are predominantly answer-first, produ

safetyarxiv-cs-cl
4 Jun 2026
Safety

Covert Influence Between Language Models

DGX agent

arXiv:2606.04071v1 Announce Type: cross Abstract: As language models increasingly consume one another's outputs, covert influence -- a phenomenon where a sender's payload (the behavioral disposition i

safetyarxiv-cs-cl
4 Jun 2026
Research

CRAFT: Cost-aware Refinement And Front-aware Tuning of Prompts

DGX agent

arXiv:2606.04661v1 Announce Type: new Abstract: Prompts tuned for accuracy often grow long, raising inference cost on every model call. The best accuracy-cost trade-off depends on the task and the bud

researcharxiv-cs-cl
4 Jun 2026
Research

Cross-Prompt Generalization in Detecting AI-Generated Fake News Using Interpretable Linguistic Features

DGX agent

arXiv:2606.04199v1 Announce Type: new Abstract: The increasing use of large language models has raised concerns about the spread of AI-generated fake news, particularly under varying prompting strateg

researcharxiv-cs-cl
4 Jun 2026
Applications

CYGNET: Cypher Gate for Neural Execution Triage and Cost Containment

DGX agent

arXiv:2606.04645v1 Announce Type: new Abstract: Language models acting as agents over knowledge graphs generate Cypher queries that fail structurally (crashing at the database) or semantically (execut

applicationsarxiv-cs-cl
4 Jun 2026
Research

Data Attribution in Large Language Models via Bidirectional Gradient Optimization

DGX agent

arXiv:2606.04928v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly deployed across diverse applications, raising critical questions for governance, accountability, and dat

researcharxiv-cs-cl
4 Jun 2026
Research

DEER: Disentangled Mixture of Experts with Instance-Adaptive Routing for Generalizable Machine-Generated Text Detection

DGX agent

arXiv:2511.01192v2 Announce Type: replace Abstract: Detecting machine-generated text has become a critical challenge amid the rapid advancement of LLMs, yet existing detectors degrade severely under d

researcharxiv-cs-cl
4 Jun 2026
Agents

Deliberate Evolution: Agentic Reasoning for Sample-Efficient Symbolic Regression with LLMs

DGX agent

arXiv:2606.04360v1 Announce Type: new Abstract: Symbolic regression (SR) discovers compact mathematical expressions from data, yet recent LLM-based evolutionary methods remain sample-inefficient becau

agentsarxiv-cs-cl
4 Jun 2026
Research

Depth-Attention: Cross-Layer Value Mixing for Language Models

DGX agent

arXiv:2606.05014v1 Announce Type: new Abstract: Self-attention selects information freely across the sequence, but across depth, Transformers merely add each layer's output to the residual stream, so

researcharxiv-cs-cl
4 Jun 2026
Model Releases

Discourse-Role Labels as Presentation-Time Variables for Context Use in Language Models

DGX agent

arXiv:2606.04109v1 Announce Type: new Abstract: Context-augmented language model systems often wrap supplied content with labels such as Reference:, Evidence:, Instruction:, Note:, or Example:, but th

model-releasesarxiv-cs-cl
4 Jun 2026
Research

Disentangling Answer Engine Optimization from Platform Growth: A Log-Based Natural Experiment on ChatGPT Referral Traffic

DGX agent

arXiv:2606.04362v1 Announce Type: cross Abstract: Large language model (LLM) 'answer engines' such as ChatGPT now send measurable referral traffic to the open web, and a practice analogous to search e

researcharxiv-cs-cl
4 Jun 2026
Model Releases

DLLG: Dynamic Logit-Level Gating of LLM Experts

DGX agent

arXiv:2606.04378v1 Announce Type: new Abstract: Leveraging multiple specialized LLMs can combine complementary strengths, but existing approaches trade adaptability for stability: routing commits prem

model-releasesarxiv-cs-cl
4 Jun 2026
Safety

DuDi: Dual-Signal Distillation with Cross-Lingual Verbalizer

DGX agent

arXiv:2606.04694v1 Announce Type: new Abstract: Small language models (SLMs) are efficient and scalable, but their multilingual capabilities degrade severely at sub-billion scales, especially for Sout

safetyarxiv-cs-cl
4 Jun 2026
Tutorials

Effective vocabulary expansion of multilingual language models for extremely low-resource languages

DGX agent

arXiv:2602.09388v2 Announce Type: replace Abstract: Multilingual pre-trained language models(mPLMs) offer significant benefits for many low-resource languages. To further expand the range of languages

tutorialsarxiv-cs-cl
4 Jun 2026
Local Ai

Efficient Reasoning on the Edge

DGX agent

arXiv:2603.16867v2 Announce Type: replace-cross Abstract: Large language models (LLMs) with chain-of-thought reasoning achieve state-of-the-art performance across complex problem-solving tasks, but th

local-aiarxiv-cs-cl
4 Jun 2026
Research

Enhancing Hallucination Detection through Noise Injection

DGX agent

arXiv:2502.03799v4 Announce Type: replace Abstract: Large Language Models (LLMs) are prone to generating plausible yet incorrect responses, known as hallucinations. Effectively detecting hallucination

researcharxiv-cs-cl
4 Jun 2026
Research

Entity Binding Failures in Speech LLM Reasoning: Diagnosis and Chain-of-Thought Intervention

DGX agent

arXiv:2606.04474v1 Announce Type: new Abstract: Speech Large Language Models (SLLMs) underperform their text counterparts on complex reasoning. We reveal that this modality gap is not a uniform cognit

researcharxiv-cs-cl
4 Jun 2026
Research

Evaluating Autoformalization Robustness via Semantically Similar Paraphrasing

DGX agent

arXiv:2511.12784v3 Announce Type: replace Abstract: Large Language Models (LLMs) have recently emerged as powerful tools for autoformalization. Despite their impressive performance, these models can s

researcharxiv-cs-cl
4 Jun 2026
Model Releases

Evaluating Large Language Models in Dynamic Clinical Decision-Making with Standardized Patient Cases

DGX agent

arXiv:2606.05112v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly proposed as clinical agents, yet static, single-turn benchmarks cannot capture how a model dynamically del

model-releasesarxiv-cs-cl
4 Jun 2026
Safety

Expert-Aware Refusal Steering

DGX agent

arXiv:2606.04160v1 Announce Type: new Abstract: Safety alignment in instruction-tuned large language models (LLMs) depends on a model's ability to reliably refuse to respond to harmful or disallowed r

safetyarxiv-cs-cl
4 Jun 2026
Model Releases

Exploring the Topology and Memory of Consensus: How LLM Agents Agree, Fragment, or Settle When Forming Conventions

DGX agent

arXiv:2606.04197v1 Announce Type: cross Abstract: How much should an LLM agent remember, and how should multi-agent systems be connected when trying to reach consensus? We show these two design choice

model-releasesarxiv-cs-cl
4 Jun 2026
Research

Fast & Faithful Function Vectors

DGX agent

arXiv:2606.05079v1 Announce Type: new Abstract: Function vectors (FVs) are task representations elicited during in-context learning that can be used to steer Large Language Models (LLMs). However, des

researcharxiv-cs-cl
4 Jun 2026
Safety

Few Tokens, Big Leverage: Preserving Safety Alignment by Constraining Safety Tokens during Fine-tuning

DGX agent

arXiv:2603.07445v2 Announce Type: replace Abstract: Large language models (LLMs) often require fine-tuning (FT) to perform well on downstream tasks, but FT can induce safety-alignment drift even when

safetyarxiv-cs-cl
4 Jun 2026
Applications

Fine-grained Fragment Retrieval in Multi-modal Long-form Dialogues

DGX agent

arXiv:2606.04591v1 Announce Type: new Abstract: With the widespread adoption of multi-modal communication platforms, long-form dialogues interleaving text and images have become increasingly common. U

applicationsarxiv-cs-cl
4 Jun 2026
Safety

GARL: Game-Theoretic Reinforcement Learning for Multi-Agent Strategic Prioritisation

DGX agent

arXiv:2606.05002v1 Announce Type: new Abstract: LLM-based multi-agent systems are increasingly used for strategic decision-making tasks. In such settings, performance depends not only on individual mo

safetyarxiv-cs-cl
4 Jun 2026
Model Releases

GENEB: Why Genomic Models Are Hard to Compare

DGX agent

arXiv:2606.04525v1 Announce Type: new Abstract: Progress in genomic foundation models is difficult to assess due to fragmented benchmarks, incompatible evaluation protocols, and task-specific reportin

model-releasesarxiv-cs-cl
4 Jun 2026
Research

GIFT: Games as Informal Training for Generalizable LLMs

DGX agent

arXiv:2601.05633v2 Announce Type: replace Abstract: Recent LLMs excel at formal tasks such as mathematical reasoning and code generation, but still struggle with broader abilities such as planning, cr

researcharxiv-cs-cl
4 Jun 2026
Safety

Global Sketch-Based Watermarking for Diffusion Language Models

DGX agent

arXiv:2606.04486v1 Announce Type: cross Abstract: Watermarking methods for language models have been studied extensively in the autoregressive setting, where tokens are generated sequentially. These w

safetyarxiv-cs-cl
4 Jun 2026
Research

GlossAssist -- A Tool to Simplify Corpus Creation and Study the Effect of NLP Models in Low-Resource Documentation Settings

DGX agent

arXiv:2606.04367v1 Announce Type: new Abstract: Interlinear glossed text (IGT) is the standard format for linguistic annotation in language documentation. Producing it manually, however, is often slow

researcharxiv-cs-cl
4 Jun 2026
Safety

GRAIL: Gradient-Reweighted Advantages for Reinforcement Learning with Verifiable Rewards

DGX agent

arXiv:2606.04889v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (e.g. GRPO) is now a common way to improve mathematical reasoning in Large Language Models (LLMs). Howeve

safetyarxiv-cs-cl
4 Jun 2026
Agents

Graph-R1: Towards Agentic GraphRAG Framework via End-to-end Reinforcement Learning

DGX agent

arXiv:2507.21892v2 Announce Type: replace Abstract: Retrieval-Augmented Generation (RAG) mitigates hallucination in LLMs by incorporating external knowledge, but relies on chunk-based retrieval that l

agentsarxiv-cs-cl
4 Jun 2026
Model Releases

High-Quality Entity Segmentation and Grounding

DGX agent

arXiv:2402.02555v2 Announce Type: replace-cross Abstract: In this work, we propose ESG, a pipeline for high-quality entity segmentation and grounding supported by a new dataset EntitySeg. At first, th

model-releasesarxiv-cs-cl
4 Jun 2026
Safety

Hybrid Adversarial Defence for Natural Language Understanding Tasks

DGX agent

arXiv:2606.04612v1 Announce Type: new Abstract: Large Language Models (LLMs) are vulnerable both to hallucination and adversarial manipulation. Although these problems are closely related, existing de

safetyarxiv-cs-cl
4 Jun 2026
Safety

Imbuing Large Language Models with Bidirectional Logic for Robust Chain Repair

DGX agent

arXiv:2606.05030v1 Announce Type: new Abstract: Autoregressive chain-of-thought (CoT) reasoning in large language models (LLMs) is fundamentally forward-directed: each step conditions only on prior to

safetyarxiv-cs-cl
4 Jun 2026
Safety

In-Context Graphical Inference

DGX agent

arXiv:2606.05042v1 Announce Type: cross Abstract: Marginal inference in discrete graphical models forces a choice between exactness and scalability: exact algorithms are intractable for high-treewidth

safetyarxiv-cs-cl
4 Jun 2026
Safety

Large Language Models in K-12 Education: Alignment with State Curriculum Standards and Student Personas

DGX agent

arXiv:2606.04846v1 Announce Type: new Abstract: As Large Language Models (LLMs) become increasingly popular in educational settings, they raise important questions about the ethical implications of th

safetyarxiv-cs-cl
4 Jun 2026
Research

LazyAttention: Efficient Retrieval-Augmented Generation with Deferred Positional Encoding

DGX agent

arXiv:2606.04302v1 Announce Type: new Abstract: Key-value (KV) caching accelerates inference of large language models (LLMs) by reusing past computations for generated tokens. Its importance becomes e

researcharxiv-cs-cl
4 Jun 2026
Model Releases

LDARNet: DNA Adaptive Representation Network with Learnable Tokenization for Genomic Modeling

DGX agent

arXiv:2606.04552v1 Announce Type: new Abstract: Genomic foundation models increasingly adopt large language model architectures, yet almost universally rely on fixed tokenization schemes such as k-mer

model-releasesarxiv-cs-cl
4 Jun 2026
Tutorials

Learning What to Learn: Stage-Specific Data Sets for SFT-then-RL in Small Language Model Reasoning

DGX agent

arXiv:2606.04466v1 Announce Type: new Abstract: Post-training Small Language Models (SLMs) for reasoning typically follows an SFT-then-RL pipeline, yet existing work rarely considers what data should

tutorialsarxiv-cs-cl
4 Jun 2026
← Previous
1…5556575859…161
Next →