AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlog
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
Hardware

Staged Factorial Screening for Budget-Constrained Micro-Pretraining

DGX agent

arXiv:2606.05186v1 Announce Type: cross Abstract: Budget-constrained micro-pretraining often requires triaging many candidate recipes on a shared accelerator before larger search budgets are spent. We

hardwarearxiv-cs-cl
5 Jun 2026
Local Ai
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Statistical Priors for Implicit Preferences: Decoupling Skill Selection as a Local Harness in Personal Agents

DGX agent

arXiv:2606.05828v1 Announce Type: cross Abstract: As Large Language Model (LLM) capabilities advance, locally deployed personal agents relying on API-based remote models and external skills have emerg

local-aiarxiv-cs-cl
5 Jun 2026
Model Releases

Statistically Reliable LLM-Based Ranking Evaluation via Prediction-Powered Inference

DGX agent

arXiv:2606.05308v1 Announce Type: cross Abstract: With PRECISE, we extended Prediction-Powered Inference to produce bias-corrected estimates of ranking evaluation metrics by combining a small human-la

model-releasesarxiv-cs-cl
5 Jun 2026
Agents

Staying with the Uncertainty: Uncertainty-Scaffolding Strategies for Artificial Moral Advisors in LLM-to-LLM Simulated Conversations

DGX agent

arXiv:2606.05890v1 Announce Type: new Abstract: LLMs are increasingly deployed as Artificial Moral Advisors (AMA) in a variety of contexts: what kind of conversational patterns should they display? In

agentsarxiv-cs-cl
5 Jun 2026
Model Releases

SubtleMemory: A Benchmark for Fine-Grained Relational Memory Discrimination in Long-Horizon AI Agents

DGX agent

arXiv:2606.05761v1 Announce Type: cross Abstract: Persistent AI assistants, such as OpenClaw, accumulate large collections of related memories over long-term interactions. As these memories grow, they

model-releasesarxiv-cs-cl
5 Jun 2026
Model Releases

TARPO: Token-Wise Latent-Explicit Reasoning via Action-Routing Policy Optimization

DGX agent

arXiv:2606.05859v1 Announce Type: new Abstract: Latent reasoning has emerged as a promising alternative to discrete Chain-of-Thought (CoT) in large language models (LLMs), enabling more expressive rea

model-releasesarxiv-cs-cl
5 Jun 2026
Local Ai

Temporal Preference Concepts and their Functions in a Large Language Model

DGX agent

arXiv:2606.05194v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly being deployed to make decisions that require trading off near-term gains against long-term consequences

local-aiarxiv-cs-cl
5 Jun 2026
Model Releases

Ten Headache Specialists versus Artificial Intelligence for Clinical Literature Summarization: A Critical Evaluation and Comparison

DGX agent

arXiv:2606.05436v1 Announce Type: cross Abstract: Summarizing the latest medical literature to guide clinical decision-making is essential for evidence-based medicine and high-quality patient care. Ye

model-releasesarxiv-cs-cl
5 Jun 2026
Model Releases

TensorBench: Benchmarking Coding Agents on a Compiler-Based Tensor Framework

DGX agent

arXiv:2606.05570v1 Announce Type: new Abstract: Repository-level coding benchmarks face a trade-off between task difficulty and evaluation reliability: tasks that challenge frontier models often invol

model-releasesarxiv-cs-cl
5 Jun 2026
Applications

The Generator-Eraser Paradox: Community Guidelines for Responsible LLM-Assisted Dialect Resource Creation

DGX agent

arXiv:2606.06004v1 Announce Type: new Abstract: Dialect resources occupy a unique position at the intersection of scientific description, cultural preservation, and computational infrastructure. Large

applicationsarxiv-cs-cl
5 Jun 2026
Model Releases

The Granularity Gap: A Multi-Dimensional Longitudinal Audit of Sycophancy in Gemini Models

DGX agent

arXiv:2606.05183v1 Announce Type: new Abstract: Large language models are increasingly deployed as high-stakes advisors, yet standard alignment benchmarks treat sycophancy as a binary failure mode. We

model-releasesarxiv-cs-cl
5 Jun 2026
Model Releases

The Mirage of Performance Gains: Why Contrastive Decoding Fails to Mitigate Object Hallucinations in MLLMs?

DGX agent

arXiv:2504.10020v4 Announce Type: replace Abstract: Contrastive decoding strategies are widely used to reduce object hallucinations in multimodal large language models (MLLMs). These methods work by c

model-releasesarxiv-cs-cl
5 Jun 2026
Applications

The Prosody of Emojis

DGX agent

arXiv:2508.00537v2 Announce Type: replace Abstract: Prosodic features such as pitch, timing, and intonation are central to spoken communication, conveying emotion, intent, and discourse structure. In

applicationsarxiv-cs-cl
5 Jun 2026
Agents

The Self-Correction Illusion: LLMs Correct Others but Not Themselves

DGX agent

arXiv:2606.05976v1 Announce Type: cross Abstract: Recent work shows that LLM agents struggle to correct errors in their own reasoning traces yet show markedly higher correction rates when identical cl

agentsarxiv-cs-cl
5 Jun 2026
Research

The Tell-Tale Norm: ell_2 Magnitude as a Signal for Reasoning Dynamics in Large Language Models

DGX agent

arXiv:2606.06188v1 Announce Type: new Abstract: Recent work has sought to understand Large Language Models (LLMs) reasoning, yet a principled, model-intrinsic signal that captures its layer-wise reaso

researcharxiv-cs-cl
5 Jun 2026
Applications

To Be Multimodal or Not to Be: Query-Adaptive Audio-Visual Person Retrieval via Active Modality Detection

DGX agent

arXiv:2606.05931v1 Announce Type: new Abstract: When retrieving a person from a video archive by voice and face, should the system be multimodal or not? In real-world broadcast archives, unlike curate

applicationsarxiv-cs-cl
5 Jun 2026
Safety

Toward Culturally Aligned LLMs through Ontology-Guided Multi-Agent Reasoning

DGX agent

arXiv:2601.21700v3 Announce Type: replace Abstract: Large Language Models (LLMs) increasingly support culturally sensitive decision making, yet often exhibit misalignment due to skewed pretraining dat

safetyarxiv-cs-cl
5 Jun 2026
Research

Towards Truly Multilingual ASR: Generalizing Code-Switching ASR to Unseen Language Pairs

DGX agent

arXiv:2606.05846v1 Announce Type: new Abstract: Automatic Speech Recognition (ASR) has become a key technology for human--AI interaction. However, code-switching ASR (CS-ASR) remains particularly chal

researcharxiv-cs-cl
5 Jun 2026
Local Ai

Trajectory Dynamics in Language Model Hidden States Predict Human Processing Costs Beyond Surprisal

DGX agent

arXiv:2606.05346v1 Announce Type: new Abstract: Human language comprehension unfolds sequentially: each word is processed in the context of those that came before, and the interpretation builds increm

local-aiarxiv-cs-cl
5 Jun 2026
Safety

UNIVID: Unified Vision-Language Model for Video Moderation

DGX agent

arXiv:2606.05748v1 Announce Type: cross Abstract: Global-scale video moderation faces a dual challenge: the need for fine-grained multi-modal reasoning and the demand for interpretable outputs to supp

safetyarxiv-cs-cl
5 Jun 2026
Agents

Unsupervised Skill Discovery for Agentic Data Analysis

DGX agent

arXiv:2606.06416v1 Announce Type: cross Abstract: Inference-time skill augmentation provides a lightweight way to improve data-analytic agents by injecting reusable procedural knowledge without updati

agentsarxiv-cs-cl
5 Jun 2026
Research

USAD 2.0: Scaling Representation Distillation for Universal Audio Understanding

DGX agent

arXiv:2606.06444v1 Announce Type: cross Abstract: Audio encoders are critical to modern audio applications as large language models (LLMs) increasingly rely on a single encoder for diverse inputs. Whi

researcharxiv-cs-cl
5 Jun 2026
Model Releases

Using Large Language Models to Support High Volume Application Review for an Undergraduate Research Program

DGX agent

arXiv:2606.05564v1 Announce Type: new Abstract: Undergraduate research programs such as the Summer Undergraduate Research Fellowship (SURF) at Purdue University receive thousands of applications every

model-releasesarxiv-cs-cl
5 Jun 2026
Safety

Value-and-Structure Alignment for Routing-Consistent Quantization of Mixture-of-Experts Models

DGX agent

arXiv:2606.05688v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) models scale foundation models efficiently by activating only a subset of experts for each token, but their large number of exp

safetyarxiv-cs-cl
5 Jun 2026
Research

Vavanagi: a Community-run Platform for Documentation of the Hula Language in Papua New Guinea

DGX agent

arXiv:2603.14210v2 Announce Type: replace Abstract: We present Vavanagi, a community-run platform for Hula (Vula'a), an Austronesian language of Papua New Guinea with approximately 10,000 speakers. Va

researcharxiv-cs-cl
5 Jun 2026
Research

What Makes Two Language Models Think Alike?

DGX agent

arXiv:2406.12620v3 Announce Type: replace Abstract: Do architectural and training differences influence the way models represent and process language? Traditional similarity metrics tell us whether tw

researcharxiv-cs-cl
5 Jun 2026
Safety

What's in a Name? Morphological Shortcuts by LLMs in Pharmacology

DGX agent

arXiv:2606.05616v1 Announce Type: new Abstract: The morphological form of a word can often give cues to its meaning, but purely relying on these mappings can lead to overgeneralization in high-stakes

safetyarxiv-cs-cl
5 Jun 2026
Safety

When AI Says It Feels

DGX agent

arXiv:2606.05734v1 Announce Type: cross Abstract: Large language models (LLMs) are generally constrained from expressing feelings through human-preference alignment in post-training processes. This po

safetyarxiv-cs-cl
5 Jun 2026
Safety

When Evidence is Sparse: Weakly Supervised Early Failure Alerting in Dialogs and LLM-Agent Trajectories

DGX agent

arXiv:2606.05414v1 Announce Type: new Abstract: Early failure alerting requires deciding, while a dialog or agent trajectory is still unfolding, whether to flag it as likely to fail. This is challengi

safetyarxiv-cs-cl
5 Jun 2026
Research

When New Generators Arrive: Lifelong Machine-Generated Text Attribution via Ridge Feature Transfer

DGX agent

arXiv:2606.05626v1 Announce Type: new Abstract: Machine-generated text (MGT) attribution aims to identify the specific generator responsible for a given text, thereby providing fine-grained evidence f

researcharxiv-cs-cl
5 Jun 2026
Research

Where does Absolute Position come from in decoder-only Transformers?

DGX agent

arXiv:2606.06160v1 Announce Type: cross Abstract: RoPE-trained transformers distinguish absolute position in their attention patterns, even though RoPE encodes only relative offsets in the inner produ

researcharxiv-cs-cl
5 Jun 2026
Model Releases

Would you still call this Dax? Novel Visual References in VLMs and Humans

DGX agent

arXiv:2606.05409v1 Announce Type: cross Abstract: Vision-language models (VLMs), like human learners, are frequently exposed to new visual concepts, but how they map novel visual references to languag

model-releasesarxiv-cs-cl
5 Jun 2026
Research

You Only Index Once: Cross-Layer Sparse Attention with Shared Routing

DGX agent

arXiv:2606.06467v1 Announce Type: new Abstract: Long-context inference in modern LLMs is increasingly constrained by decoding efficiency, especially in reasoning-heavy settings where models generate l

researcharxiv-cs-cl
5 Jun 2026
Model Releases

YouZhi: Towards High-Concurrency Financial LLMs via Adaptive GQA-to-MLA Transition

DGX agent

arXiv:2606.05868v1 Announce Type: new Abstract: Large language models (LLMs) drive significant financial innovations, yet their high-concurrency deployment is severely bottlenecked by KV cache memory

model-releasesarxiv-cs-cl
5 Jun 2026
Research

A French Corpus Annotated for Multiword Expressions with Adverbial Function

DGX agent

arXiv:2606.04828v1 Announce Type: new Abstract: This paper presents a French corpus annotated for multiword expressions (MWEs) with adverbial function. This corpus is designed for investigation on inf

researcharxiv-cs-cl
4 Jun 2026
Model Releases

A Systematic Evaluation of Positional Bias in Multi-Video Summarization with MLLMs

DGX agent

arXiv:2606.04596v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) are increasingly used for video understanding, yet their reliability under multi-video inputs remains poorly un

model-releasesarxiv-cs-cl
4 Jun 2026
Research

ACAT: A Collaborative Platform for Efficient Aspect-Based Sentiment Dataset Annotation

DGX agent

arXiv:2606.04189v1 Announce Type: new Abstract: Aspect-Based Sentiment Analysis (ABSA) requires high-quality datasets to train reliable models. However, existing annotation tools treat output as flat

researcharxiv-cs-cl
4 Jun 2026
Model Releases

Activation-Based Active Learning for In-Context Learning: Challenges and Insights

DGX agent

arXiv:2606.05134v1 Announce Type: new Abstract: Deep active learning has previously been explored for LLM in-context sample selection, but not with methods that utilise recent advances in understandin

model-releasesarxiv-cs-cl
4 Jun 2026
Safety

Adaptive Information Control for Search-Augmented LLM Reasoning

DGX agent

arXiv:2602.01672v2 Announce Type: replace Abstract: Search-augmented reasoning agents interleave multi-step reasoning with external retrieval, but uncontrolled retrieval can introduce redundant eviden

safetyarxiv-cs-cl
4 Jun 2026
Model Releases

Agent Planning Benchmark: A Diagnostic Framework for Planning Capabilities in LLM Agents

DGX agent

arXiv:2606.04874v1 Announce Type: new Abstract: Planning is central to LLM agents: before acting, an agent must decompose goals, select tools, reason over constraints, and decide when a task is infeas

model-releasesarxiv-cs-cl
4 Jun 2026
Model Releases

AUDDT: A Unified Benchmark Toolkit for Audio and Speech Deepfake Detectors

DGX agent

arXiv:2509.21597v2 Announce Type: replace-cross Abstract: With the prevalence of artificial intelligence (AI)-generated content, such as audio deepfakes, a large body of recent work has focused on dev

model-releasesarxiv-cs-cl
4 Jun 2026
Research

Automated Lexical Coverage for Language Learning: From General to Specialized Word Lists

DGX agent

arXiv:2512.15552v2 Announce Type: replace Abstract: A General Service List (GSL) is a commonly used resource for language learners to identify important English words. Traditional GSL creation is reso

researcharxiv-cs-cl
4 Jun 2026
Applications

BEATS: Bootstrapping E-commerce Attribute Taxonomies for Search through Iterative Human-AI Collaboration

DGX agent

arXiv:2606.04909v1 Announce Type: cross Abstract: E-commerce platforms in emerging markets often operate with underdeveloped product catalogs that contain only category taxonomies but lack structured

applicationsarxiv-cs-cl
4 Jun 2026
Model Releases

Benchmarking Living-Screen-Native GUI Agents on Short-Video Platforms

DGX agent

arXiv:2606.04701v1 Announce Type: cross Abstract: GUI agents today assume a static screen, where the world is frozen between two actions. However, real interfaces such as short-video applications viol

model-releasesarxiv-cs-cl
4 Jun 2026
Agents

Beyond Correctness: Rewarding Faithful Reasoning in Retrieval-Augmented Generation

DGX agent

arXiv:2510.13272v3 Announce Type: replace Abstract: Inspired by the success of reinforcement learning (RL) in Large Language Model (LLM) training for domains like math and code, recent work has begun

agentsarxiv-cs-cl
4 Jun 2026
Model Releases

Beyond Retrieval: Learning Compact User Representations for Scalable LLM Personalization

DGX agent

arXiv:2606.04547v1 Announce Type: cross Abstract: Personalizing large language models requires adapting model behavior to individual users while preserving robustness and deployment-scale efficiency.

model-releasesarxiv-cs-cl
4 Jun 2026
Research

Beyond Text Following: Repairable Arbitration Reversals in Audio-Language Models

DGX agent

arXiv:2606.05161v1 Announce Type: cross Abstract: Audio-language models (ALMs) often follow text that conflicts with audio, even when the audio evidence is clear. This raises a basic question: is the

researcharxiv-cs-cl
4 Jun 2026
Research

Boosting Self-Consistency with Ranking

DGX agent

arXiv:2606.05054v1 Announce Type: new Abstract: Self-consistency improves large language models by sampling multiple reasoning paths and selecting the most frequent answer, but majority voting often f

researcharxiv-cs-cl
4 Jun 2026
← Previous
1…5455565758…161
Next →