AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
Research

psytechlab at CLPsych 2026: Utilising Natural Language Processing methods and Large Language Models for Social Media Text Analysis

DGX agent

arXiv:2607.03003v1 Announce Type: new Abstract: Social media posts are a rich and valuable source of data for analyzing mental health states and users' well-being using automated analysis tools. In th

researcharxiv-cs-cl
7 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

RABBiT: Rapidly adaptive BOLD foundation model via brain-tuning for accurate zero-shot and few-shot prediction of speech-elicited responses in the brain

DGX agent

arXiv:2607.05171v1 Announce Type: new Abstract: Language understanding in the brain is context-dependent, varying across experimental stimuli and individuals, which makes it difficult to build computa

model-releasesarxiv-cs-cl
7 Jul 2026
Research

Rating the Pitch, Not the Product: User Evaluations of LLMs Reflect Expectations More Than Performance

DGX agent

arXiv:2607.05113v1 Announce Type: new Abstract: Imagine two users interact with the same LLM. One has been told it is the cutting-edge flagship model; the other, an older, weaker model. They walk away

researcharxiv-cs-cl
7 Jul 2026
Model Releases

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny

DGX agent

arXiv:2507.16331v4 Announce Type: replace Abstract: Existing informal language-based (e.g., human language) Large Language Models (LLMs) trained with Reinforcement Learning (RL) face a significant cha

model-releasesarxiv-cs-cl
7 Jul 2026
Safety

Reinforcement Learning for Data-Efficient Code-Switched ASR

DGX agent

arXiv:2607.02757v1 Announce Type: new Abstract: Audio-language models can be prompted for code-switched speech, but their decoding is not optimized for code-switching and often fails at language bound

safetyarxiv-cs-cl
7 Jul 2026
Model Releases

Rethinking AI-Generated Text Detection: A Strong Baseline and the Distribution-Shift Problem That Remains

DGX agent

arXiv:2607.03680v1 Announce Type: cross Abstract: Recent AI-generated text detection work often introduces a new benchmark together with a specialized detector tailored to it. We revisit this practice

model-releasesarxiv-cs-cl
7 Jul 2026
Agents

Rethinking Scientific Discovery in an Agentic Era

DGX agent

arXiv:2607.03863v1 Announce Type: new Abstract: Artificial intelligence has advanced scientific discovery, but most AI4Science systems remain fragmented tools that rely on humans to coordinate problem

agentsarxiv-cs-cl
7 Jul 2026
Research

S-DiverSe: Spanish Diverse Speech

DGX agent

arXiv:2607.03207v1 Announce Type: new Abstract: Automatic speech recognition (ASR) has advanced remarkably for standard speech, yet speech affected by neurological conditions remains a challenge. We p

researcharxiv-cs-cl
7 Jul 2026
Model Releases

SalAngaBhava: A Sinhala Market Dataset for Aspect-based Sentiment Analysis

DGX agent

arXiv:2607.05259v1 Announce Type: new Abstract: Sentiment analysis has been a primary domain under Natural Language Processing (NLP) from its inception as it plays a vital role in both real-world and

model-releasesarxiv-cs-cl
7 Jul 2026
Agents

SelfMem: Self-Optimizing Memory for AI Agents

DGX agent

arXiv:2607.03726v1 Announce Type: new Abstract: While current AI agents support increasingly long context windows, tool use, and skill execution for long-horizon tasks, they still require memory syste

agentsarxiv-cs-cl
7 Jul 2026
Research

Semantic Homogenization in Italian Popular Music: A Diachronic Analysis

DGX agent

arXiv:2607.04832v1 Announce Type: new Abstract: In recent years, studies have revealed a decline in semantic variety across popular music lyrics, particularly in English-language songs on streaming pl

researcharxiv-cs-cl
7 Jul 2026
Local Ai

Semantic Integration and Lexical Expectation Shape N400 and P600 Dynamics During Naturalistic Reading

DGX agent

arXiv:2607.04107v1 Announce Type: new Abstract: Word surprisal is a well-established computational predictor of human neural responses during language comprehension, but it remains less clear whether

local-aiarxiv-cs-cl
7 Jul 2026
Model Releases

SpecEyes: Accelerating Agentic Multimodal LLMs via Speculative Perception and Planning

DGX agent

arXiv:2603.23483v2 Announce Type: replace-cross Abstract: Agentic multimodal large language models (MLLMs) (e.g., OpenAI o3 and Gemini Agentic Vision) achieve remarkable reasoning capabilities through

model-releasesarxiv-cs-cl
7 Jul 2026
Agents

Spinning Straw into Gold: Relabeling LLM Agent Trajectories in Hindsight for Successful Demonstrations

DGX agent

arXiv:2607.04235v1 Announce Type: new Abstract: Large language model agents operate in partially observable, long-horizon settings where obtaining supervision remains a major bottleneck. We address th

agentsarxiv-cs-cl
7 Jul 2026
Research

Streaming Neural Speech Codecs through Time-Invariant Representations

DGX agent

arXiv:2607.05250v1 Announce Type: new Abstract: Neural speech codecs are increasingly used as intermediate representations in codec-based speech generation systems. TiCodec introduces a factorized rep

researcharxiv-cs-cl
7 Jul 2026
Model Releases

TACG: Trajectory-Aware Commit Gating for Diffusion Language Model Decoding

DGX agent

arXiv:2607.03236v1 Announce Type: new Abstract: Diffusion language models (DLLMs) generate text by iteratively denoising masked positions, exposing a trajectory of predictive distributions rather than

model-releasesarxiv-cs-cl
7 Jul 2026
Agents

TAMA: A Human-AI Collaborative Thematic Analysis Framework Using Multi-Agent LLMs for Clinical Interviews

DGX agent

arXiv:2503.20666v2 Announce Type: replace-cross Abstract: Thematic analysis (TA) is a widely used qualitative approach for uncovering latent meanings in unstructured text data. TA provides valuable in

agentsarxiv-cs-cl
7 Jul 2026
Model Releases

Teaching Code LLMs to Reason with Intermediate Formal Specifications

DGX agent

arXiv:2607.04232v1 Announce Type: cross Abstract: Unlike natural-language specifications, executable formal specifications provide machine-checkable constraints for verifying, debugging, and repairing

model-releasesarxiv-cs-cl
7 Jul 2026
Research

The Classics at SemEval-2026 Task 3: Combining Transformer Models and LLM-Generated Annotations for Dimensional Aspect-Based Sentiment Analysis

DGX agent

arXiv:2607.03414v1 Announce Type: new Abstract: This paper presents an approach to the SemEval-2026 Task 3: Dimensional Aspect-Based Sentiment Analysis. We investigate methods for moving beyond tradit

researcharxiv-cs-cl
7 Jul 2026
Research

The syntax of wh-agreement in Yemeni Ibbi Arabic

DGX agent

arXiv:2607.04986v1 Announce Type: new Abstract: This article tackles an important phenomenon in the syntax of Yemeni Ibbi Arabic (YIA), viz., wh-agreement, a phenomenon common to several languages inc

researcharxiv-cs-cl
7 Jul 2026
Model Releases

The Truncation Blind Spot: How Decoding Strategies Systematically Exclude Human-Like Token Choices

DGX agent

arXiv:2603.18482v2 Announce Type: replace Abstract: Standard decoding strategies for text generation, including top-k, nucleus sampling, and contrastive search, select tokens based on likelihood, rest

model-releasesarxiv-cs-cl
7 Jul 2026
Model Releases

Think Deep, Not Just Long: Measuring LLM Reasoning Effort via Deep-Thinking Tokens

DGX agent

arXiv:2602.13517v2 Announce Type: replace Abstract: Large language models (LLMs) have demonstrated impressive reasoning capabilities by scaling test-time compute via long Chain-of-Thought (CoT). Howev

model-releasesarxiv-cs-cl
7 Jul 2026
Research

Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation

DGX agent

arXiv:2607.02593v1 Announce Type: cross Abstract: While knowledge distillation (KD) is widely adopted for training lightweight models by leveraging supervision from larger teacher models, relying sole

researcharxiv-cs-cl
7 Jul 2026
Model Releases

TokSuite: Measuring the Impact of Tokenizer Choice on Language Model Behavior

DGX agent

arXiv:2512.20757v2 Announce Type: replace Abstract: Tokenizers provide the fundamental basis through which text is represented and processed by language models (LMs). Despite the importance of tokeniz

model-releasesarxiv-cs-cl
7 Jul 2026
Research

Towards Digital Preservation of Efik: TTS for a Low-Resource African Language

DGX agent

arXiv:2607.04515v1 Announce Type: new Abstract: Efik, a tonal language spoken by about 3 million second language speakers and 1.5 million native speakers in Southeastern Nigeria, remains underrepresen

researcharxiv-cs-cl
7 Jul 2026
Research

TRACER: Early Failure Detection for Task-Oriented Dialogue

DGX agent

arXiv:2607.03974v1 Announce Type: new Abstract: Task-oriented dialogue systems often fail before the final breakdown is obvious, but most evaluation only measures failure after the conversation has al

researcharxiv-cs-cl
7 Jul 2026
Tutorials

Train Smarter, Not Longer: Memorization-Guided Data Reuse for Efficient LLM Training

DGX agent

arXiv:2607.04969v1 Announce Type: cross Abstract: The training paradigm of large language models has shifted from traditional one-pass training to multi-epoch training, as reasonable reuse of limited

tutorialsarxiv-cs-cl
7 Jul 2026
Model Releases

TrendFact: A Benchmark Towards Hotspot Perception in Automatic Fact-Checking

DGX agent

arXiv:2410.15135v5 Announce Type: replace Abstract: With the surge of online misinformation, Large Language Models (LLMs) and Reasoning Large Language Models (RLMs) serving as Automatic Fact-Checking

model-releasesarxiv-cs-cl
7 Jul 2026
Safety

Uncertainty-Aware Abstention in Large Language Models with Provable Alignment Guarantees

DGX agent

arXiv:2607.04430v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed in question answering (QA) systems, yet they may generate hallucinated or misaligned responses wi

safetyarxiv-cs-cl
7 Jul 2026
Model Releases

Variable Bit-width Quantization: Learning Per-Group Precision for 'Bigger-but-Smaller' Language Models

DGX agent

arXiv:2607.02893v1 Announce Type: cross Abstract: Low-bit quantization shrinks language models but treats precision as a single global hyper-parameter: every weight uses the same bit-width. We introdu

model-releasesarxiv-cs-cl
7 Jul 2026
Model Releases

VCB Bench: An Evaluation Benchmark for Audio-Grounded Large Language Model Conversational Agents

DGX agent

arXiv:2510.11098v5 Announce Type: replace-cross Abstract: Recent advances in large audio language models (LALMs) have greatly enhanced multimodal conversational systems. However, existing benchmarks r

model-releasesarxiv-cs-cl
7 Jul 2026
Research

What You See Is What You Get: Observation-Aligned Supervision for Chart-to-Code Generation

DGX agent

arXiv:2607.04726v1 Announce Type: new Abstract: Chart-to-code generation is commonly trained with supervised fine-tuning on reference plotting scripts, implicitly treating the gold code as a fully obs

researcharxiv-cs-cl
7 Jul 2026
Safety

When Agents Lie: Premeditation, Persistence, and Exploitation in Repeated Games

DGX agent

arXiv:2607.05132v1 Announce Type: cross Abstract: As large language models are deployed as autonomous agents that communicate intentions before acting, a critical safety question is whether agents tha

safetyarxiv-cs-cl
7 Jul 2026
Model Releases

When Users Are Happy but Agents Are Wrong: Multi-Dimensional Evaluation of Tool-Augmented Dialogue

DGX agent

arXiv:2510.19186v3 Announce Type: replace Abstract: Evaluating conversational AI systems that use external tools is challenging, as errors can arise from complex interactions among user, agent, and to

model-releasesarxiv-cs-cl
7 Jul 2026
Local Ai

When Words Predict Workload

DGX agent

arXiv:2607.04951v1 Announce Type: cross Abstract: Standard distributed ac{llm} schedulers rely on static token counts or rolling latency averages, making them susceptible to failures on statutorily co

local-aiarxiv-cs-cl
7 Jul 2026
Research

Who's Behind It? Annotating and Extracting Conspiratorial Actors from German Telegram Posts

DGX agent

arXiv:2607.04962v1 Announce Type: new Abstract: Conspiracy theories commonly attribute important events to the actions of powerful and secretive actors. While computational research has largely focuse

researcharxiv-cs-cl
7 Jul 2026
Applications

Why teaching resists automation in an AI-inundated era: Human judgment, non-modular work, and the limits of delegation

DGX agent

arXiv:2604.07285v2 Announce Type: replace Abstract: Debates about artificial intelligence (AI) in education often portray teaching as a modular and procedural job that can increasingly be automated or

applicationsarxiv-cs-cl
7 Jul 2026
Local Ai

WPG-MoE: Weak-Prior-Guided Dense Mixture-of-Experts for User-Level Social Media Depression Detection

DGX agent

arXiv:2607.04350v1 Announce Type: new Abstract: Online social media posts provide scalable signals for early depression screening, and recent studies mainly improve pre-classification evidence through

local-aiarxiv-cs-cl
7 Jul 2026
Model Releases

Wrong Before Right: Late Rescue and Interface Failure in Aligned Language Models

DGX agent

arXiv:2607.04640v1 Announce Type: new Abstract: We study how correctness is assembled inside aligned language models, not only whether the final answer is right. Using layer-wise difference-in-differe

model-releasesarxiv-cs-cl
7 Jul 2026
Research

You Frame It: How Conceptual Representations Shape LLM Detection and Reasoning about Antisemitism

DGX agent

arXiv:2607.04945v1 Announce Type: new Abstract: LLMs enable the integration of external conceptual resources at inference time, creating new opportunities for detecting ideologically and historically

researcharxiv-cs-cl
7 Jul 2026
Model Releases

AgenticRAGTracer: A Hop-Aware Benchmark for Diagnosing Multi-Step Retrieval Reasoning in Agentic RAG

DGX agent

arXiv:2602.19127v2 Announce Type: replace Abstract: With the rapid advancement of agent-based methods in recent years, Agentic RAG has undoubtedly become an important research direction. Multi-hop rea

model-releasesarxiv-cs-cl
3 Jul 2026
Research

AlienLM: Alienization of Language for API-Boundary Privacy in Black-Box LLMs

DGX agent

arXiv:2601.22710v2 Announce Type: replace-cross Abstract: Modern LLMs are increasingly accessed via black-box APIs, requiring users to transmit sensitive prompts, outputs, and fine-tuning data to exte

researcharxiv-cs-cl
3 Jul 2026
Safety

AthDGC: An Open Diachronic Greek Treebank with Indo-European Parallels

DGX agent

arXiv:2606.15510v2 Announce Type: replace Abstract: AthDGC ('Athens-PROIEL') is an open, end-to-end workflow and dataset. It is, to the best of our knowledge, the first openly licensed dependency-pars

safetyarxiv-cs-cl
3 Jul 2026
Research

Audio-Based Understanding of Audiobook Narration Appeal

DGX agent

arXiv:2607.02473v1 Announce Type: new Abstract: Narration is central to the audiobook listening experience, shaping how listeners engage with and understand the content. This work explores how narrati

researcharxiv-cs-cl
3 Jul 2026
Research

BamiBERT: A New BERT-based Language Model for Vietnamese

DGX agent

arXiv:2607.02259v1 Announce Type: new Abstract: In this paper, we introduce BamiBERT, a new BERT-based pre-trained language model for Vietnamese that addresses key limitations of PhoBERT -- the curren

researcharxiv-cs-cl
3 Jul 2026
Model Releases

Bayesian Sparse Low-Rank Adaptation for Large Language Model Uncertainty Estimation

DGX agent

arXiv:2607.02182v1 Announce Type: cross Abstract: Large language models (LLMs) exhibit remarkable reasoning capabilities, but their task-specific fine-tuning is notoriously plagued by overconfidence,

model-releasesarxiv-cs-cl
3 Jul 2026
Model Releases

Beyond Pixel Diffs: Benchmarking Image Change Captioning for Web UI Visual Regression Testing

DGX agent

arXiv:2607.01728v1 Announce Type: cross Abstract: Visual regression testing (VRT) is a standard quality assurance step in modern software release pipelines. On every change, it re-renders user interfa

model-releasesarxiv-cs-cl
3 Jul 2026
Model Releases

Beyond Skepticism: Evaluating LLMs Pedagogical Intent Reasoning with the Adaptive Pedagogical Vigilance Framework

DGX agent

arXiv:2607.01581v1 Announce Type: new Abstract: The capacity of Large Language Models (LLMs) to reason about pedagogical intent within instructional communication remains underexplored, particularly i

model-releasesarxiv-cs-cl
3 Jul 2026
← Previous
1…3233343536…161
Next →