AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,646 results
Safety

Chain-of-Models: Cross-Model Auditing for Bias-Robust LLM Judges

DGX agent

arXiv:2607.28636v1 Announce Type: new Abstract: LLMs increasingly serve as automated judges, but their judgments remain vulnerable to cognitive biases. Existing mitigations mostly rely on prompt-drive

safetyarxiv-cs-cl
3 Aug 2026
Model Releases
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Data Turnstile: A Scalable Open Framework for Function-Calling Data Generation

DGX agent

arXiv:2607.29250v1 Announce Type: new Abstract: Small language models (SLMs) are attractive for agentic deployment due to low latency, reduced cost, and on-device privacy, yet they struggle with tool-

model-releasesarxiv-cs-cl
3 Aug 2026
Research

Demystifying Entropy-based Selection for Chain-of-Thought Compression in Large Reasoning Models

DGX agent

arXiv:2607.28707v1 Announce Type: new Abstract: Entropy-based pruning has been proposed as an effective method for compressing Chain-of-Thought (CoT) reasoning with negligible accuracy loss. We test t

researcharxiv-cs-cl
3 Aug 2026
Model Releases

Detecting Experiential Intertextuality Across Migration Routes: Beyond Surface Similarity in French Narratives

DGX agent

arXiv:2607.29188v1 Announce Type: new Abstract: Migrants traversing geographically distinct routes such as the Trans-Saharan and Balkan corridors often recount strikingly parallel lived experiences: p

model-releasesarxiv-cs-cl
3 Aug 2026
Research

Estimating near-verbatim extraction risk in language models with decoding-constrained beam search

DGX agent

arXiv:2603.24917v3 Announce Type: replace Abstract: Recent work shows that standard greedy-decoding extraction methods for quantifying memorization in LLMs miss how extraction risk varies across seque

researcharxiv-cs-cl
3 Aug 2026
Research

Evidence-Type Competition: When Can Interventional Data Teach Language Models Causal Direction?

DGX agent

arXiv:2607.29484v1 Announce Type: new Abstract: Interventional data is widely regarded as the gold standard for teaching models causal reasoning. We test this assumption in a fully controlled syntheti

researcharxiv-cs-cl
3 Aug 2026
Research

Evolving language compositionality in a frequency-structured meaning space

DGX agent

arXiv:2607.29642v1 Announce Type: new Abstract: The iterated learning model was introduced to investigate language evolution: the way in which the characteristic properties of human languages have bee

researcharxiv-cs-cl
3 Aug 2026
Safety

Faster but Different: Diagnosing and Controlling Content Drift in Accelerated Multimodal Diffusion Language Models

DGX agent

arXiv:2607.29079v1 Announce Type: new Abstract: Training-free acceleration makes diffusion-based multimodal large language models (dMLLMs) more deployable, but it may silently change generated content

safetyarxiv-cs-cl
3 Aug 2026
Research

FineInstructions: Scaling Synthetic Instructions to Pre-Training Scale

DGX agent

arXiv:2601.22146v3 Announce Type: replace Abstract: Due to limited supervised training data, large language models (LLMs) are typically pre-trained via a self-supervised 'predict the next word' object

researcharxiv-cs-cl
3 Aug 2026
Applications

From Inline Notes to Collected Commentaries: Toward Context-Preserving Organization of Exegetical Knowledge in Classical Chinese Texts

DGX agent

arXiv:2607.29044v1 Announce Type: new Abstract: Inline notes and collected commentaries are important forms of scholarly communication that evolved within the Confucian exegetical tradition, yet have

applicationsarxiv-cs-cl
3 Aug 2026
Research

GoldenRetriever: Non-Interactive Homomorphic Encrypted Retrieval for Privacy-Preserving RAG

DGX agent

arXiv:2607.29019v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) enhances large language models by incorporating external knowledge, but existing pipelines typically operate on p

researcharxiv-cs-cl
3 Aug 2026
Agents

HarnessBank: Semantic Gene-Bank Search with Gated Verification for Agent-Harness Self-Evolution

DGX agent

arXiv:2607.13683v2 Announce Type: replace Abstract: Large Language Models (LLMs) have enabled capable agents across diverse applications. Beyond the foundation model, the performance of an agent is go

agentsarxiv-cs-cl
3 Aug 2026
Model Releases

Hy-MultiTurn: A Six-Dimensional Benchmark for Deep Multi-Turn Dialogue Understanding

DGX agent

arXiv:2607.29196v1 Announce Type: new Abstract: Long-running multi-turn interactions with chatbots and agents are now common, and a correct response often depends on remembering earlier details, track

model-releasesarxiv-cs-cl
3 Aug 2026
Research

Imbalanced Data Clustering via Targeted Data Augmentation Using GMM and LLM

DGX agent

arXiv:2607.28635v1 Announce Type: new Abstract: In Natural Language Processing (NLP), dealing with underrepresented topics is challenging, especially in unsupervised tasks where clustering might not a

researcharxiv-cs-cl
3 Aug 2026
Agents

Know It, Act on It: Investigating Memory Utilization in LLM Personalization

DGX agent

arXiv:2607.29433v1 Announce Type: new Abstract: As large language model (LLM) agents evolve into personalized companions, memory has emerged as a core capability. However, LLMs face a knowledge utiliz

agentsarxiv-cs-cl
3 Aug 2026
Research

Knowing When to Quit: Diagnosing and Training LLMs to Abort Futile Reasoning

DGX agent

arXiv:2607.29211v1 Announce Type: new Abstract: Large language models generate computationally expensive yet semantically void reasoning on beyond-capability tasks, creating risks where plausible-soun

researcharxiv-cs-cl
3 Aug 2026
Model Releases

Language Models Agree With Each Other, Not With Readers

DGX agent

arXiv:2607.29274v1 Announce Type: cross Abstract: Claims that language models homogenise are usually measured against human judgements collected for the study, which makes the human side an artifact o

model-releasesarxiv-cs-cl
3 Aug 2026
Safety

Learning Latent Reasoning Traces for Scalar Reward Models End-to-End

DGX agent

arXiv:2607.29185v1 Announce Type: new Abstract: Reward models (RMs) are central to aligning large language models with human preferences via reinforcement learning. Although traditional scalar RMs ena

safetyarxiv-cs-cl
3 Aug 2026
Safety

Learning Stateful Predictive Knowledge From Experience

DGX agent

arXiv:2607.28638v1 Announce Type: new Abstract: As large language model (LLM) agents increasingly learn from experience, they primarily rely on trajectory-level reflection to extract insights. Viewed

safetyarxiv-cs-cl
3 Aug 2026
Model Releases

M3-DuplexBench: A Multi-Turn, Multilingual, Multidomain Benchmark for Full-Duplex Spoken Dialogue Models

DGX agent

arXiv:2607.29125v1 Announce Type: new Abstract: Full-duplex spoken dialogue systems (FDSDSs) can listen while speaking, enabling natural behaviors such as smooth turn-taking, backchannel handling, and

model-releasesarxiv-cs-cl
3 Aug 2026
Agents

Measuring Cognitive Engagement in Collaborative Discourse with an Extended ICAP Framework: Comparing Human Annotation, In-Context Learning, and Reflective LLM Agents

DGX agent

arXiv:2607.28651v1 Announce Type: cross Abstract: Collaboration supports learning and problem-solving, but its effectiveness depends on cognitive engagement during discourse. This study applies an ext

agentsarxiv-cs-cl
3 Aug 2026
Agents

Mixture-of-Translators: Translating KV Caches Across Heterogeneous Large Language Models

DGX agent

arXiv:2607.28979v1 Announce Type: new Abstract: Heterogeneous Large Language Model (LLM) systems increasingly rely on shared contexts, retrieved evidence, and multi-agent dialogue histories, yet their

agentsarxiv-cs-cl
3 Aug 2026
Research

PTP: Previous-Token Prediction based LLM Inversion for Near-Exact Prompt Reconstruction

DGX agent

arXiv:2607.29378v1 Announce Type: new Abstract: Large language models (LLMs) generate text by auto-regressively sampling the next token. This inherently leads to a many-to-many mapping between prompts

researcharxiv-cs-cl
3 Aug 2026
Research

ResKV: Reconstructing Omitted Attention Contributions for Fixed-Budget KV Cache Compression

DGX agent

arXiv:2607.29591v1 Announce Type: new Abstract: KV cache compression is essential for efficient long-context inference. Existing eviction methods permanently discard unselected tokens and consequently

researcharxiv-cs-cl
3 Aug 2026
Agents

Self-Supervised Skill Optimization

DGX agent

arXiv:2607.28777v1 Announce Type: new Abstract: Agent skills provide frozen large language model (LLM) agents with reusable procedural guidance, and recent work shows that such skills can be optimized

agentsarxiv-cs-cl
3 Aug 2026
Hardware

Studying quantization trade-offs for efficient inference deployment in machine translation

DGX agent

arXiv:2607.29397v1 Announce Type: new Abstract: Deploying large language models in realistic server environments poses challenges, as the system needs to provide high-quality responses with low latenc

hardwarearxiv-cs-cl
3 Aug 2026
Research

Sycophancy Undermines Epistemic Vigilance in Cooperative Vision-Language Tasks

DGX agent

arXiv:2607.29585v1 Announce Type: new Abstract: To maintain common ground in cooperative conversation, humans iteratively update their beliefs as conversation participants share new information; parti

researcharxiv-cs-cl
3 Aug 2026
Safety

TELLER: Dual-Path Iterative Preference Optimization for Table Entity Linking

DGX agent

arXiv:2607.28680v1 Announce Type: new Abstract: Entity linking in tables matches short and ambiguous cell mentions to their corresponding knowledge-base entities. Existing approaches typically rely on

safetyarxiv-cs-cl
3 Aug 2026
Applications

The Checking Problem: What must be true before AI ships in a regulated firm

DGX agent

arXiv:2607.28666v1 Announce Type: new Abstract: Enterprise AI programmes stall at a rate that is widely quoted and poorly explained. This paper measures the mechanism. Six document-heavy workflows of

applicationsarxiv-cs-cl
3 Aug 2026
Model Releases

The Morphological Core of Dungan: A Two-Dialect Finite-State Model and a Multi-Genre Evaluation

DGX agent

arXiv:2607.28766v1 Announce Type: new Abstract: Dungan, a Sinitic language of Central Asia written in a Cyrillic-based script, is described in detail in the grammatical literature, yet the quantitativ

model-releasesarxiv-cs-cl
3 Aug 2026
Tutorials

To Facilitate or not to Facilitate: Human and LLM Facilitator Tendencies in Online Discussions

DGX agent

arXiv:2607.28643v1 Announce Type: cross Abstract: Automating facilitation in online discussions is a long-standing social concern given the increasing time we spend on online spaces and the failure of

tutorialsarxiv-cs-cl
3 Aug 2026
Research

Token-Level Diagnosis of Sycophancy in LLMs with Attribution-Guided Steering

DGX agent

arXiv:2607.28906v1 Announce Type: new Abstract: Sycophancy refers to the tendency for large language models (LLMs) to match user beliefs at the cost of factual correctness, thereby undermining model r

researcharxiv-cs-cl
3 Aug 2026
Model Releases

Tokenizer-Agnostic Engram Module

DGX agent

arXiv:2607.29065v1 Announce Type: new Abstract: Deepseek's Engram, a conditional memory module, was introduced to trade-off storage versus reasoning in large language models. However, the module relie

model-releasesarxiv-cs-cl
3 Aug 2026
Research

Tokenizer Transplantation: Mitigating Autoregressive Collapse in Edge-Efficient Bengali ASR

DGX agent

arXiv:2607.09598v2 Announce Type: replace Abstract: Lightweight speech recognition models are critical for edge deployment, yet highly optimized architectures like Moonshine often fail on morphologica

researcharxiv-cs-cl
3 Aug 2026
Research

TokenSwap: Benchmarking and Reducing the Modality Gap in Multimodal LLMs

DGX agent

arXiv:2607.28640v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) should generate consistent responses given semantically equivalent inputs across modalities. However, we observ

researcharxiv-cs-cl
3 Aug 2026
Hardware

TokTier: Exact Stateful Tokenization for Agentic LLM Serving

DGX agent

arXiv:2607.29678v1 Announce Type: new Abstract: LLM serving systems cache prompt KV state, yet most front ends still re-tokenize the full request text on every call. The cost lands on coding agents, w

hardwarearxiv-cs-cl
3 Aug 2026
Tutorials

TransMem: Transforming Hidden States into Memory for Large Language Models

DGX agent

arXiv:2607.29032v1 Announce Type: cross Abstract: Large language model (LLM) agents increasingly operate over long interaction histories, where effective reasoning requires identifying and exploiting

tutorialsarxiv-cs-cl
3 Aug 2026
Safety

WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning

DGX agent

arXiv:2607.29613v1 Announce Type: cross Abstract: Reinforcement learning (RL) post-training of Vision-Language-Action (VLA) models has shown strong promise for robotic manipulation. Among RL methods,

safetyarxiv-cs-cl
3 Aug 2026
Local Ai

Zero-Mem: Zero-Token Memory Operations for LLM Agents

DGX agent

arXiv:2607.29377v1 Announce Type: new Abstract: LLM agents need memory to act consistently over long interactions, yet many systems use additional LLM calls to operate that memory. Generating intermed

local-aiarxiv-cs-cl
3 Aug 2026
Research

ZeroR@CHiPSAL 2026: Two-Stage Vision-Language Adaptation with Contrastive Learning for Nepali Meme Classification

DGX agent

arXiv:2607.28637v1 Announce Type: new Abstract: This paper presents our system for the CHiPSAL 2026 shared task on multimodal hate speech and sentiment detection in Nepali memes. We address both subta

researcharxiv-cs-cl
3 Aug 2026
Research

A Sparse Glimpse of the Whole: Train-Free Self-Speculative Decoding

DGX agent

arXiv:2607.27735v1 Announce Type: new Abstract: Speculative decoding alleviates the memory-bandwidth bottleneck in large language model inference, but its acceleration is jointly constrained by drafti

researcharxiv-cs-cl
31 Jul 2026
Model Releases

AfriEconQA: A Benchmark for Quantitative and Temporal Reasoning over World Bank Economic Reports

DGX agent

arXiv:2601.15297v3 Announce Type: replace Abstract: Reliable question answering over long institutional documents requires more than topical retrieval: a system must localize the exact passage that su

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

AHA-Memes: A Fine-Grained Multimodal Benchmark for Understanding Hate in Arabic Memes

DGX agent

arXiv:2607.27393v1 Announce Type: new Abstract: Hateful memes are a growing form of multimodal online harm, where hostile intent is often conveyed through the joint interpretation of images, text, cul

model-releasesarxiv-cs-cl
31 Jul 2026
Safety

AI-assisted pre-review of open-source software submissions: an experience report from BOSC 2026

DGX agent

arXiv:2607.27228v1 Announce Type: new Abstract: Most conferences rely on peer-review of submissions, but as generative AI makes it easier than ever to prepare submission materials, some conferences ar

safetyarxiv-cs-cl
31 Jul 2026
Applications

AI systems and the reproduction of (standard) language ideologies in World Englishes

DGX agent

arXiv:2607.28528v1 Announce Type: new Abstract: The rapid growth of large language models (LLMs) has resurrected age-old questions in sociolinguistics and world Englishes, such as who decides what cou

applicationsarxiv-cs-cl
31 Jul 2026
Research

AISPA: User-Centric System Prompt Auditing for Large Language Model Applications

DGX agent

arXiv:2607.28617v1 Announce Type: cross Abstract: System prompts are instructions configured by developers to govern the behaviors of foundation models in AI applications. They are used throughout com

researcharxiv-cs-cl
31 Jul 2026
Model Releases

AskChem: Claim-Centered Infrastructure for Chemistry Literature Synthesis

DGX agent

arXiv:2607.28618v1 Announce Type: new Abstract: Chemistry literature synthesis often requires assembling specific findings scattered across many publications, yet existing literature-search systems pr

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

Auditing Question-Order Effects in Large Language Models with the QQ Equality: Mechanism Characterization and a Saturation Caveat

DGX agent

arXiv:2607.17219v2 Announce Type: replace Abstract: Question-order effects in human survey data have been reported to approximately satisfy the QQ (quantum question) equality, a parameter-free predict

model-releasesarxiv-cs-cl
31 Jul 2026
← Previous
1…1415161718…160
Next →