AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlog
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,646 results
Applications

Building Open-Retrieval Conversational Question Answering Systems by Generating Synthetic Data and Decontextualizing User Questions

DGX agent

arXiv:2507.04884v2 Announce Type: replace Abstract: We consider open-retrieval conversational question answering (OR-CONVQA), an extension of question answering where system responses need to be (i) a

applicationsarxiv-cs-cl
7 Aug 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Agents

CalibForge: Adversarial Solver Calibration for Scaling Learnable Terminal Tasks

DGX agent

arXiv:2608.06352v1 Announce Type: cross Abstract: Training terminal agents requires executable and verifiable tasks that are not merely solvable, but appropriately challenging for learning. Executable

agentsarxiv-cs-cl
7 Aug 2026
Model Releases

Causal Episodic Memory for Feedback-Driven Agent Repair

DGX agent

arXiv:2608.05906v1 Announce Type: new Abstract: LLM agents that repair failures often discard successful corrections, forcing later episodes to rediscover similar solutions. We study whether finalized

model-releasesarxiv-cs-cl
7 Aug 2026
Safety

Clinical Communication Processing with Models Trained on LLM-Generated Synthetic Data: A Structured Survey and Novel Application Case Studies

DGX agent

arXiv:2608.05993v1 Announce Type: new Abstract: Much clinical value is conveyed not through structured records but through communication: exchanges in which patients describe symptoms, clinicians reas

safetyarxiv-cs-cl
7 Aug 2026
Model Releases

CNM-BERT: A Drop-In Structural Embedding for Chinese Characters via Ideographic Description Sequences

DGX agent

arXiv:2608.05167v1 Announce Type: new Abstract: Token-based encoders like BERT treat Chinese characters as atomic identifiers, ignoring their recursive orthographic structure. Consequently, models rel

model-releasesarxiv-cs-cl
7 Aug 2026
Model Releases

Conditional Cognitive Biases in LLMs: How Biased User Turns Modulate In-Context Reasoning

DGX agent

arXiv:2608.05166v1 Announce Type: new Abstract: We present an evaluation of cognitive bias expression in state-of-the-art instruction-tuned LLMs under realistic multi-turn interaction settings. Our wo

model-releasesarxiv-cs-cl
7 Aug 2026
Research

Constraint-First Reasoning: A Training-Free Protocol for Exploiting Answer-Space Constraints in Mathematical Problem Solving

DGX agent

arXiv:2608.05254v1 Announce Type: new Abstract: Large language models can derive a plausible mathematical object yet still violate explicit requirements--for example, by omitting a modular reduction,

researcharxiv-cs-cl
7 Aug 2026
Model Releases

ConWriter: Transition-Constrained Stateful Long-Form Story Generation with Lightweight Neuro-Symbolic Consistency Control

DGX agent

arXiv:2608.05169v1 Announce Type: new Abstract: Long-form story generation requires models to preserve narrative consistency across extended contexts, yet existing prompting-based methods often accumu

model-releasesarxiv-cs-cl
7 Aug 2026
Research

CPC-CMS: Cognitive Pairwise Comparison Classification Model Selection Framework for Document-level Sentiment Analysis

DGX agent

arXiv:2507.14022v2 Announce Type: replace Abstract: This study proposes the Cognitive Pairwise Comparison Classification Model Selection (CPC-CMS) framework for document-level sentiment analysis. The

researcharxiv-cs-cl
7 Aug 2026
Model Releases

Cross-Architecture Steering Transfer in Language Models: A Systematic Empirical Study

DGX agent

arXiv:2608.05164v1 Announce Type: new Abstract: Independently trained large language models may develop shared internal representations of semantic concepts despite architectural differences -- but wh

model-releasesarxiv-cs-cl
7 Aug 2026
Research

DBLAST: Dependent Block Drafting for Stochastic Speculative Decoding

DGX agent

arXiv:2608.05448v1 Announce Type: new Abstract: Speculative decoding accelerates large language models' inference by using a lightweight drafter to propose multiple future tokens and a target model to

researcharxiv-cs-cl
7 Aug 2026
Safety

Decolonizing Linguistic Policies in Automated Speech Recognition: A Framework for Cross-Culturally Competent Speech AI

DGX agent

arXiv:2608.06141v1 Announce Type: new Abstract: This paper focuses on automatic speech recognition (ASR) and ASR-mediated voice interfaces that shape access to public services, healthcare, and educati

safetyarxiv-cs-cl
7 Aug 2026
Research

Decomposed Entailment for Factuality Checking and Hallucination Detection

DGX agent

arXiv:2608.05823v1 Announce Type: new Abstract: The reliability of Large Language Models (LLMs) is often compromised by factual inconsistencies, including hallucinations---cases where generated conten

researcharxiv-cs-cl
7 Aug 2026
Safety

Different Perturbations, Different Mechanisms: Understanding Continued Pre-training for Zero-Shot Dialect Robustness

DGX agent

arXiv:2608.05510v1 Announce Type: new Abstract: Dialectal variation remains a major challenge for multilingual language models. Perturbation-based continued pre-training (CPT) has emerged as a promisi

safetyarxiv-cs-cl
7 Aug 2026
Local Ai

EdgeXpert: An Edge Device for Memory-Efficient LLM Inference with Mixture-of-Experts and Speculative Decoding

DGX agent

arXiv:2608.05303v1 Announce Type: cross Abstract: On-device deployment of Large Language Models (LLMs) has become essential for personalized edge applications. A primary bottleneck is external memory

local-aiarxiv-cs-cl
7 Aug 2026
Model Releases

Enhancing Social Intelligence in LLMs with Hierarchical Reasoning and Utterance-Level Goal Rewarding

DGX agent

arXiv:2608.05832v1 Announce Type: new Abstract: Large language models (LLMs) excel in structured tasks but struggle with dynamic social interactions, where success requires long-term goal coordination

model-releasesarxiv-cs-cl
7 Aug 2026
Model Releases

EpiBench: Can LLMs Understand Epitopes for Antibody Drug Discovery?

DGX agent

arXiv:2608.06022v1 Announce Type: new Abstract: Epitopes determine where antibodies bind antigens and shape downstream therapeutic properties such as functional blockade and escape resistance, making

model-releasesarxiv-cs-cl
7 Aug 2026
Model Releases

Evidence Lock Before Commitment: A Frozen Interface Degrades LLM-as-Judge Evaluation

DGX agent

arXiv:2608.05353v1 Announce Type: new Abstract: LLM judges are often asked to extract criteria and evidence before choosing between candidate answers. This workflow assumes that the intermediate recor

model-releasesarxiv-cs-cl
7 Aug 2026
Safety

EvoHarness-RL: Learning Self-Evolving Runtime Harness for Long-Horizon LLM Agents

DGX agent

arXiv:2608.05446v1 Announce Type: cross Abstract: Long-horizon LLM agents increasingly rely on external execution support to maintain state, track progress, invoke tools, verify outcomes, and reuse ex

safetyarxiv-cs-cl
7 Aug 2026
Research

Example-Guided Prompting for Document-Level Text Simplification

DGX agent

arXiv:2608.05447v1 Announce Type: new Abstract: Document-level text simplification requires large language models (LLMs) to rewrite complex documents while preserving meaning, readability, and discour

researcharxiv-cs-cl
7 Aug 2026
Applications

FOCUS: Decoupling Expert Personas in LLMs to Enhance Domain Expert Capabilities

DGX agent

arXiv:2608.05611v1 Announce Type: new Abstract: Large Language Models (LLMs) can exhibit diverse personas, and activating expert personas has been shown to improve domain expertise and task accuracy.

applicationsarxiv-cs-cl
7 Aug 2026
Model Releases

From Sports to Safety: Benchmarking Proactive Risk Inference in MLLMs

DGX agent

arXiv:2608.05560v1 Announce Type: cross Abstract: Timely anticipation of physical hazards is essential for real-world safety, yet existing MLLM evaluations focus on harmful content or general risks, l

model-releasesarxiv-cs-cl
7 Aug 2026
Safety

GenGA: Editable and Data-Grounded Graphical Abstract Generation for Academic Papers

DGX agent

arXiv:2608.05478v1 Announce Type: cross Abstract: Graphical Abstracts (GAs) visually summarize the key findings of academic papers, playing a crucial role in facilitating the understanding of research

safetyarxiv-cs-cl
7 Aug 2026
Tutorials

How to Recognize New Words: A Comparison Between Context Biasing Methods and Speech LLMs

DGX agent

arXiv:2608.05759v1 Announce Type: new Abstract: Recognizing new and rare words - named entities, acronyms, domain specific special words, and other items scarce in training data - remains a key challe

tutorialsarxiv-cs-cl
7 Aug 2026
Model Releases

Human-Like Anaphor Resolution in Large Language Models

DGX agent

arXiv:2608.05630v1 Announce Type: new Abstract: Anaphors are expressions that refer to other expressions, called antecedents. The process of connecting the two is called resolution. Cognitive science

model-releasesarxiv-cs-cl
7 Aug 2026
Research

Integrating Human Linguistic Insights into AI: Theory-Driven Representation for Multilingual Text-to-Speech

DGX agent

arXiv:2204.07228v2 Announce Type: replace Abstract: This paper explores the integration of human linguistic insights into multilingual text-to-speech (TTS) systems by evaluating the Featurally Undersp

researcharxiv-cs-cl
7 Aug 2026
Model Releases

LangChoiceBench: Measuring and Explaining Programming-Language Choice in LLMs

DGX agent

arXiv:2608.06041v1 Announce Type: cross Abstract: Large language models (LLMs) have been shown to exhibit strong Python preferences when generating project-level code, but there is currently no system

model-releasesarxiv-cs-cl
7 Aug 2026
Research

LELA: an LLM-based Entity Linking Approach with Zero-Shot Domain Adaptation

DGX agent

arXiv:2601.05192v2 Announce Type: replace Abstract: Entity linking (mapping ambiguous mentions in text to entities in a knowledge base) is a foundational step in tasks such as knowledge graph construc

researcharxiv-cs-cl
7 Aug 2026
Model Releases

M^3R-Bench: A Unified Benchmark for Evidence-Grounded Multimodal Metaphor Understanding

DGX agent

arXiv:2608.05817v1 Announce Type: new Abstract: Metaphor enables the understanding of abstract concepts through cross-domain mappings while conveying affective attitudes. In multimodal scenarios, visu

model-releasesarxiv-cs-cl
7 Aug 2026
Safety

Mapping Patient-Perceived Physician Traits from Nationwide Online Reviews with LLMs

DGX agent

arXiv:2510.03997v2 Announce Type: replace Abstract: Understanding how patients perceive their physicians is essential to improving trust, communication, and satisfaction. Patients increasingly consult

safetyarxiv-cs-cl
7 Aug 2026
Tutorials

Mapping Similarity Spaces across Embedding Models with Synthetic Query Probing

DGX agent

arXiv:2608.05857v1 Announce Type: new Abstract: Retrieval-Augmented Generation systems rely on similarity scores to retrieve relevant content, yet scores are not directly comparable across embedding m

tutorialsarxiv-cs-cl
7 Aug 2026
Safety

Mitigating Scoring Bias in LLM-as-a-Judge via Random Number Generation

DGX agent

arXiv:2608.05726v1 Announce Type: new Abstract: Large Language Models (LLMs) are often used as evaluators of text quality, known as LLM-as-a-Judge, which can outperform conventional automatic evaluati

safetyarxiv-cs-cl
7 Aug 2026
Model Releases

MoCA: Implicit Social Context Analysis

DGX agent

arXiv:2608.05825v1 Announce Type: new Abstract: Human social communication, such as affection and intent, is often conveyed in highly implicit ways, where underlying meanings are expressed through ind

model-releasesarxiv-cs-cl
7 Aug 2026
Safety

Mood Matters: How Syntactic Sensitivity Undermines Safety Alignment

DGX agent

arXiv:2608.05409v1 Announce Type: new Abstract: Large language models typically undergo post-training to align them with safety policies but there exist many sophisticated jailbreaks that sidestep est

safetyarxiv-cs-cl
7 Aug 2026
Model Releases

NeSy-RAG: Neuro-Symbolic RAG for Explainable Question Answering

DGX agent

arXiv:2608.06292v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) improves question answering by grounding large language models (LLMs) in external knowledge such as text corpora. H

model-releasesarxiv-cs-cl
7 Aug 2026
Safety

On-Policy Delta Distillation for Multilingual Math Reasoning

DGX agent

arXiv:2608.05802v1 Announce Type: new Abstract: On-Policy Distillation (OPD) is emerging as a promising alternative to reinforcement learning for LLM post-training, yet its effectiveness in multilingu

safetyarxiv-cs-cl
7 Aug 2026
Applications

Persona-Pruner: Sculpting Lightweight Models for Role-Playing

DGX agent

arXiv:2606.14695v2 Announce Type: replace-cross Abstract: Language Models (LMs) have shown remarkable potential as role-playing chatbots, delivering consistent, stylized interactions when given a spec

applicationsarxiv-cs-cl
7 Aug 2026
Safety

PolyAlign: Conditional Human-Distribution Alignment

DGX agent

arXiv:2606.13227v2 Announce Type: replace Abstract: Post-training methods such as supervised fine-tuning (SFT) and preference optimization typically align language models toward a single global assist

safetyarxiv-cs-cl
7 Aug 2026
Model Releases

PoolBench: A Benchmark for Pooling Strategies in Concept Representation Evaluation for Decoder-Only LLMs

DGX agent

arXiv:2608.05162v1 Announce Type: new Abstract: Pooling is a consequential but under-examined design choice in decoder-only concept representation work: practitioners must collapse token-level hidden

model-releasesarxiv-cs-cl
7 Aug 2026
Research

Predicting Social Media User Actions: A Hybrid Approach for Common and Rare Behavior Prediction on Bluesky

DGX agent

arXiv:2511.17241v2 Announce Type: replace Abstract: Understanding and predicting user behavior on social media platforms is crucial for content recommendation and platform design. While existing appro

researcharxiv-cs-cl
7 Aug 2026
Agents

Predicting Task Difficulty Without Rollouts

DGX agent

arXiv:2608.05797v1 Announce Type: cross Abstract: Task difficulty dictates an agent's likelihood of success, and estimating it without rollouts means forecasting this directly from a task description

agentsarxiv-cs-cl
7 Aug 2026
Research

QEvict: Recoverable Quantized KV Eviction for Attention-Drift-Robust Long-Context Decoding

DGX agent

arXiv:2608.05326v1 Announce Type: cross Abstract: Autoregressive large language model inference is increasingly constrained by the memory footprint of the Key-Value (KV) cache. A dominant line of work

researcharxiv-cs-cl
7 Aug 2026
Research

Reasoning Errors Have a Region and a Direction in the Residual-Stream Trajectory of LLMs

DGX agent

arXiv:2608.05660v1 Announce Type: cross Abstract: As language models are increasingly used for tasks that require verifiable reasoning, reliably distinguishing sound reasoning from flawed reasoning ha

researcharxiv-cs-cl
7 Aug 2026
Research

RIG-RoPE: Relation- and Instance-Gated Rotary Positional Encoding with Duration-Aware Temporal Coordinates

DGX agent

arXiv:2608.05154v1 Announce Type: new Abstract: Rotary positional encoding (RoPE) is a core component of modern language models and has been extended to multimodal LLMs through multidimensional varian

researcharxiv-cs-cl
7 Aug 2026
Model Releases

Robust Native Language Identification through Agentic Decomposition

DGX agent

arXiv:2509.16666v2 Announce Type: replace Abstract: Large language models (LLMs) often achieve high performance in native language identification (NLI) benchmarks by leveraging superficial contextual

model-releasesarxiv-cs-cl
7 Aug 2026
Agents

Routing Is Least Learnable Where It Is Most Valuable: Bounds on Representation Routing for Web Agents

DGX agent

arXiv:2608.06171v1 Announce Type: new Abstract: Web agents observe a browser through text, pixels, or both, and the choice is usually fixed once for all tasks. We measure six observation modes across

agentsarxiv-cs-cl
7 Aug 2026
Safety

RP-OPSD: Reasoning-Pivot-Guided On-Policy Self-Distillation for Multilingual Reasoning Transfer

DGX agent

arXiv:2608.06347v1 Announce Type: new Abstract: Multilingual reasoning transfer is crucial for extending reasoning capabilities of large language models (LLMs) beyond high-resource languages. On-polic

safetyarxiv-cs-cl
7 Aug 2026
Research

RRC: Unlocking Generative Reward Models in LLM Reinforcement Learning via Ranking-Based Reward Construction

DGX agent

arXiv:2608.06310v1 Announce Type: cross Abstract: Recent advances in reward modeling show a paradigm shift from discriminative reward models to generative reward models. However, despite their strong

researcharxiv-cs-cl
7 Aug 2026
← Previous
1…45678…160
Next →