AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlog
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,759 results
Research

Ineffectiveness for Search and Undecidability of PCSP Meta-Problems

DGX agent

arXiv:2504.04639v4 Announce Type: replace-cross Abstract: It is an open question whether the search and decision versions of promise CSPs are equivalent. Most known algorithms for PCSPs solve only the

researcharxiv-cs-cl
26 May 2026
Model Releases
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Inference Time Optimization with Confidence Dynamics

DGX agent

arXiv:2605.25244v1 Announce Type: new Abstract: Inference time optimization techniques, such as repeated sampling, have significantly advanced the reasoning capabilities of Large Language Models (LLMs

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

InfiFPO: Implicit Model Fusion via Preference Optimization in Large Language Models

DGX agent

arXiv:2505.13878v3 Announce Type: replace-cross Abstract: Model fusion combines multiple Large Language Models (LLMs) with different strengths into a more powerful, integrated model through lightweigh

model-releasesarxiv-cs-cl
26 May 2026
Research

Is Inference Mediated by Distinct Semantic Structures in LLMs? A Mechanistic Interpretation

DGX agent

arXiv:2605.25520v1 Announce Type: new Abstract: Predicting a label correctly does not necessarily require representing the operation that produces it. Transformer representations are known to carry la

researcharxiv-cs-cl
26 May 2026
Agents

Iterate Until Retrieved: Factual Nugget Optimization for Discoverable Continual Corrections in Agentic RAG

DGX agent

arXiv:2605.25641v1 Announce Type: new Abstract: Agentic retrieval-augmented generation (RAG) systems in complex B2B (business-to-business) settings may often receive free-form response feedback. Rathe

agentsarxiv-cs-cl
26 May 2026
Applications

Know You Before You Speak: User-State Modeling for LLM Personalization in Multi-Turn Conversation

DGX agent

arXiv:2605.24647v1 Announce Type: new Abstract: Personalized dialogue requires more than recalling explicit user histories: systems also need to infer hidden user states that evolve through interactio

applicationsarxiv-cs-cl
26 May 2026
Research

Knowing but Not Showing: LLMs Recognize Ambiguity but Rarely Ask Clarifying Questions

DGX agent

arXiv:2605.25284v1 Announce Type: new Abstract: User queries are often underspecified and may admit multiple valid interpretations. Rather than silently making assumptions about the user's intent, a h

researcharxiv-cs-cl
26 May 2026
Research

Large Language Model Selection with Limited Annotations

DGX agent

arXiv:2605.24981v1 Announce Type: new Abstract: Choosing a Large Language Model (LLM) for a given task requires comparing many strong candidates, yet standard evaluation relies on costly annotations o

researcharxiv-cs-cl
26 May 2026
Research

Lean Formalization of Generalization Error Bound by Rademacher Complexity and Dudley's Entropy Integral

DGX agent

arXiv:2503.19605v5 Announce Type: replace-cross Abstract: Understanding and certifying the generalization performance of machine learning algorithms -- i.e. obtaining theoretical estimates of the test

researcharxiv-cs-cl
26 May 2026
Safety

Learning Concepts, Not Tokens: Self-Supervised Semantic Alignment for Language Models

DGX agent

arXiv:2603.29123v2 Announce Type: replace Abstract: The next-token prediction (NTP) objective trains language models to predict a single token at each step, even though many continuations can express

safetyarxiv-cs-cl
26 May 2026
Safety

Learning to Route Languages for Multilingual Policy Optimization

DGX agent

arXiv:2605.25360v1 Announce Type: new Abstract: Large language models~(LLMs) are trained on heterogeneous multilingual corpora, yet existing policy optimization methods often implicitly restrict each

safetyarxiv-cs-cl
26 May 2026
Model Releases

Llamion Technical Report

DGX agent

arXiv:2605.25676v1 Announce Type: new Abstract: We release Llamion, a family of 14B-parameter open-weight language models obtained by transforming Orion-14B into the standardized Llama-family architec

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

LLM-as-a-Reviewer: Benchmarking Their Ability, Divergence, and Prompt Injection Resistance as Paper Reviewers

DGX agent

arXiv:2605.25415v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used in academic peer review, yet their reliability, alignment with human judgment, and robustness to adve

model-releasesarxiv-cs-cl
26 May 2026
Local Ai

Lngram: N-gram Conditional Memory in Latent Space

DGX agent

arXiv:2605.24869v1 Announce Type: new Abstract: Sequence modeling requires both compositional reasoning and local static knowledge retrieval, yet standard Transformers handle both through dense comput

local-aiarxiv-cs-cl
26 May 2026
Safety

Locality Matters for Training-Free Audio Token Compression in Audio-Language Models

DGX agent

arXiv:2605.25179v1 Announce Type: new Abstract: Audio-language models (ALMs) are increasingly used for audio captioning, question answering, and open-ended audio understanding, but their inference cos

safetyarxiv-cs-cl
26 May 2026
Safety

MAGIC: Multimodal Alignment & Grounding-aware Instruction Coreset for Vision-Language Models

DGX agent

arXiv:2605.26004v1 Announce Type: cross Abstract: Instruction tuning of large vision-language models (LVLMs) increasingly depends on massive multimodal corpora, yet these datasets contain samples with

safetyarxiv-cs-cl
26 May 2026
Research

Mapping the Schedule x Bit-Width Boundary in Sub-100M Quantisation-Aware Training

DGX agent

arXiv:2605.25966v1 Announce Type: cross Abstract: We test whether the optimal learning-rate schedule depends on bit-width during from-initialisation quantisation-aware training (QAT) for sub-100M deco

researcharxiv-cs-cl
26 May 2026
Safety

MATO: Multi-objective Personalized Alignment with Test-time Optimization for Large Language Models

DGX agent

arXiv:2605.25342v1 Announce Type: new Abstract: Aligning large language models (LLMs) with diverse and multifaceted user preferences is a fundamental challenge in personalized AI systems. Existing mul

safetyarxiv-cs-cl
26 May 2026
Model Releases

MindAlign: Bridging EEG, Vision, and Language for Zero-Shot Visual Decoding

DGX agent

arXiv:2605.24523v1 Announce Type: cross Abstract: Visual decoding from brain signals is a key challenge at the intersection of computer vision and neuroscience, requiring methods that bridge neural re

model-releasesarxiv-cs-cl
26 May 2026
Safety

Mitigating Hallucinations in Healthcare LLMs with Granular Fact-Checking and Domain-Specific Adaptation

DGX agent

arXiv:2512.16189v3 Announce Type: replace Abstract: In healthcare, it is essential for any LLM-generated output to be reliable and accurate, particularly in cases involving decision-making and patient

safetyarxiv-cs-cl
26 May 2026
Research

Mitigating Provenance-Role Collapse in Long-Term Agents via Typed Memory Representation

DGX agent

arXiv:2605.25869v1 Announce Type: new Abstract: Long-term memory is essential for persistent LLM agents, yet prevailing architectures store historical interactions as unstructured, flat text. This unc

researcharxiv-cs-cl
26 May 2026
Model Releases

MMSI-Bench: A Benchmark for Multi-Image Spatial Intelligence

DGX agent

arXiv:2505.23764v3 Announce Type: replace-cross Abstract: Spatial intelligence is essential for multimodal large language models (MLLMs) operating in the complex physical world. Existing benchmarks, h

model-releasesarxiv-cs-cl
26 May 2026
Agents

Multi-Persona Debate System for Automated Scientific Hypothesis Generation

DGX agent

arXiv:2605.23917v1 Announce Type: new Abstract: Modern scientific discovery is bottlenecked not by data scarcity, but by the inability to synthesize fragmented knowledge into actionable hypotheses. Th

agentsarxiv-cs-cl
26 May 2026
Model Releases

MultiHaluDet: Multilingual Hallucination Detection via LLM Hidden State Probing

DGX agent

arXiv:2605.24919v1 Announce Type: new Abstract: Hallucinations in Large Language Models (LLMs) represent a critical barrier to their reliable deployment, a vulnerability heavily exacerbated in non-Eng

model-releasesarxiv-cs-cl
26 May 2026
Research

Multilingual Phonological Feature Recognition with Self-Supervised Speech Models

DGX agent

arXiv:2605.25596v1 Announce Type: new Abstract: Phonological features provide a language-general and linguistically grounded representation of speech. We present PhonoQ-2.0, a multilingual frame-level

researcharxiv-cs-cl
26 May 2026
Model Releases

Neural Router: Semantic Content Matching for Agentic AI

DGX agent

arXiv:2605.25701v1 Announce Type: cross Abstract: Large language models (LLMs) can serve as the semantic-matching engine of a content-based publish/subscribe broker for agentic AI across the edge-clou

model-releasesarxiv-cs-cl
26 May 2026
Research

NITP: Next Implicit Token Prediction for LLM Pre-training

DGX agent

arXiv:2605.24956v1 Announce Type: new Abstract: Standard next-token prediction (NTP) supervises language models solely through discrete labels in the output logit space. We argue that this sparse one-

researcharxiv-cs-cl
26 May 2026
Research

On the Limits of Model Merging for Multilinguality in Pre-Training

DGX agent

arXiv:2605.25846v1 Announce Type: new Abstract: Endowing models with consistent multilingual performance can be achieved by mixing pre-training data, or post-training approaches such as language-speci

researcharxiv-cs-cl
26 May 2026
Safety

Optimizing Token Choice for Code Watermarking: An RL Approach

DGX agent

arXiv:2508.11925v3 Announce Type: replace-cross Abstract: Protecting intellectual property on LLM-generated code necessitates effective watermarking systems that can operate within code's highly struc

safetyarxiv-cs-cl
26 May 2026
Model Releases

Overview of the PsyDefDetect Shared Task at BioNLP 2026: Detecting Levels of Psychological Defense Mechanisms in Supportive Conversations

DGX agent

arXiv:2605.24907v1 Announce Type: new Abstract: We present an overview of PsyDefDetect, the shared task on detecting levels of psychological defense mechanisms in emotional support dialogues, co-locat

model-releasesarxiv-cs-cl
26 May 2026
Research

P1SCO: Social Dimensions from a Perspectivist Lens

DGX agent

arXiv:2605.25312v1 Announce Type: new Abstract: We introduce P1SCO, a dataset of social media comments collected from three distinct platforms, annotated according to ten social dimensions to capture

researcharxiv-cs-cl
26 May 2026
Safety

Peak-Then-Collapse and the Four Interface Channels of Knowledge-Graph Tool Use

DGX agent

arXiv:2605.26037v1 Announce Type: new Abstract: We test the standard RLVR tool-use recipe -- GRPO on Qwen2.5-7B-Instruct -- on a deliberately minimal knowledge-graph tool API: four Freebase navigation

safetyarxiv-cs-cl
26 May 2026
Research

PerSoMed: A Large-Scale Balanced Dataset for Persian Social Media Text Classification

DGX agent

arXiv:2602.19333v2 Announce Type: replace Abstract: This research introduces the first large-scale, well-balanced Persian social media text classification dataset, specifically designed to address the

researcharxiv-cs-cl
26 May 2026
Agents

Persuasion Should be Double-Blind: A Multi-Domain Dialogue Dataset With Faithfulness Based on Causal Theory of Mind

DGX agent

arXiv:2502.21297v2 Announce Type: replace Abstract: Persuasive dialogue is central to human communication, yet existing datasets often rely on a single language model generating both roles, producing

agentsarxiv-cs-cl
26 May 2026
Research

Phonetic Modeling of Dialectal Variation in Vietnamese Speech

DGX agent

arXiv:2605.24451v1 Announce Type: new Abstract: Vietnamese exhibits substantial dialectal phonetic variation across Northern, Central, and Southern regions, where identical lexical items may be realiz

researcharxiv-cs-cl
26 May 2026
Safety

PolyGnosis 2.0: Enhancing LLM Reasoning via Agentic Harness Engineering for Polymarket and OSINT Insight Extraction

DGX agent

arXiv:2605.25958v1 Announce Type: new Abstract: This paper introduces PolyGnosis 2.0, a pioneering multi-agent architecture designed to extract predictive intelligence by synthesizing Polymarket anoma

safetyarxiv-cs-cl
26 May 2026
Model Releases

PolySAE: Modeling Feature Interactions in Sparse Autoencoders via Polynomial Decoding

DGX agent

arXiv:2602.01322v2 Announce Type: replace-cross Abstract: Sparse autoencoders (SAEs) interpret neural network representations by decomposing activations into sparse combinations of dictionary atoms. H

model-releasesarxiv-cs-cl
26 May 2026
Research

PowLU: An Activation Function for Stable Pre-Training of LLMs

DGX agent

arXiv:2605.25704v1 Announce Type: new Abstract: In contemporary large language models (LLMs), the swish-gated linear unit (SwiGLU) activation function is widely adopted to regulate the information flo

researcharxiv-cs-cl
26 May 2026
Local Ai

Prefix Teach, Suffix Fade: Local Teachability Collapse in Strong-to-Weak On-Policy Distillation

DGX agent

arXiv:2605.13643v2 Announce Type: replace Abstract: On-policy distillation (OPD) trains a student model on its own rollouts using dense feedback from a stronger teacher. Prior literature suggests that

local-aiarxiv-cs-cl
26 May 2026
Applications

Prism: A Plug-in Reproducible Infrastructure for Scalable Multimodal Continual Instruction Tuning

DGX agent

arXiv:2605.26110v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) achieve versatility by reformulating diverse tasks into a unified instruction-following framework via instruc

applicationsarxiv-cs-cl
26 May 2026
Research

Proactive for Uncertainty: Cause-Aware Error Diagnosis and Interactive Clarification for Spoken Dialogue Systems

DGX agent

arXiv:2605.25404v1 Announce Type: new Abstract: Cascaded Automatic Speech Recognition -- Large Language Model (ASR-LLM) pipelines remain popular for industrial Spoken Dialogue Systems (SDS), primarily

researcharxiv-cs-cl
26 May 2026
Research

Probability Distributions Computed by Autoregressive Transformers

DGX agent

arXiv:2510.27118v4 Announce Type: replace Abstract: Most expressivity results for transformers treat them as language recognizers -- devices that accept or reject strings -- rather than as they are us

researcharxiv-cs-cl
26 May 2026
Model Releases

Quantifying the Impact of Translation Errors on Multilingual LLM Evaluation

DGX agent

arXiv:2605.24904v1 Announce Type: new Abstract: Machine-translated benchmarks are widely used to assess the multilingual capabilities of large language models (LLMs), yet translation errors in these b

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

QUEST: Training Frontier Deep Research Agents with Fully Synthetic Tasks

DGX agent

arXiv:2605.24218v1 Announce Type: new Abstract: Deep research agents extend the role of search engines from retrieving keyword-matched pages to synthesizing knowledge, fundamentally changing how human

model-releasesarxiv-cs-cl
26 May 2026
Research

Re-defining Humor Data Objects for AI Humor Research

DGX agent

arXiv:2605.25171v1 Announce Type: new Abstract: In most existing AI humor research, humor was treated as either 'present' or 'not present.' We explore the concept of humor as a social interaction with

researcharxiv-cs-cl
26 May 2026
Safety

Reading, Not Thinking: Understanding and Bridging the Modality Gap When Text Becomes Pixels in Multimodal LLMs

DGX agent

arXiv:2603.09095v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) can process text presented as images, yet they often perform worse than when the same content is provided a

safetyarxiv-cs-cl
26 May 2026
Safety

Reinforcement Learning from Denoising Feedback

DGX agent

arXiv:2605.25638v1 Announce Type: new Abstract: Policy loss estimation remains a fundamental and long-standing challenge in reinforcement learning (RL) for diffusion language models (dLLMs). We introd

safetyarxiv-cs-cl
26 May 2026
Research

Repeated Sequences Reveal Gaps between Large Language Models and Natural Language

DGX agent

arXiv:2605.24850v1 Announce Type: new Abstract: Evaluating whether large language models (LLMs) capture the structure of natural language beyond local fluency remains an open challenge. Existing evalu

researcharxiv-cs-cl
26 May 2026
← Previous
1…7778798081…162
Next →