AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,646 results
Safety

RL-PLUS: Countering Capability Boundary Collapse of LLMs in Reinforcement Learning with Hybrid-policy Optimization

DGX agent

arXiv:2508.00222v5 Announce Type: replace-cross Abstract: Reinforcement Learning with Verifiable Reward (RLVR) has significantly advanced the complex reasoning abilities of Large Language Models (LLMs

safetyarxiv-cs-cl
16 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Robust Reward Modeling for Large Language Models via Causal Decomposition

DGX agent

arXiv:2604.13833v1 Announce Type: new Abstract: Reward models are central to aligning large language models, yet they often overfit to spurious cues such as response length and overly agreeable tone.

model-releasesarxiv-cs-cl
16 Apr 2026
Research

Saber: An Efficient Sampling with Adaptive Acceleration and Backtracking Enhanced Remasking for Diffusion Language Model

DGX agent

arXiv:2510.18165v2 Announce Type: replace-cross Abstract: Diffusion language models (DLMs) are emerging as a powerful and promising alternative to the dominant autoregressive paradigm, offering inhere

researcharxiv-cs-cl
16 Apr 2026
Model Releases

Scaling Test-Time Compute to Achieve IOI Gold Medal with Open-Weight Models

DGX agent

arXiv:2510.14232v2 Announce Type: replace-cross Abstract: Competitive programming has become a rigorous benchmark for evaluating the reasoning and problem-solving capabilities of large language models

model-releasesarxiv-cs-cl
16 Apr 2026
Research

Social media polarization during conflict: Insights from an ideological stance dataset on Israel-Palestine Reddit comments

DGX agent

arXiv:2502.00414v2 Announce Type: replace Abstract: In politically sensitive scenarios like wars, social media serves as a platform for polarized discourse and expressions of strong ideological stance

researcharxiv-cs-cl
16 Apr 2026
Research

Sparse or Dense? A Mechanistic Estimation of Computation Density in Transformer-based LLMs

DGX agent

arXiv:2601.22795v2 Announce Type: replace Abstract: Transformer-based large language models (LLMs) are comprised of billions of parameters arranged in deep and wide computational graphs. Several studi

researcharxiv-cs-cl
16 Apr 2026
Model Releases

SpatialEvo: Self-Evolving Spatial Intelligence via Deterministic Geometric Environments

DGX agent

arXiv:2604.14144v1 Announce Type: cross Abstract: Spatial reasoning over three-dimensional scenes is a core capability for embodied intelligence, yet continuous model improvement remains bottlenecked

model-releasesarxiv-cs-cl
16 Apr 2026
Safety

SPG: Sandwiched Policy Gradient for Masked Diffusion Language Models

DGX agent

arXiv:2510.09541v3 Announce Type: replace Abstract: Diffusion large language models (dLLMs) are emerging as an efficient alternative to autoregressive models due to their ability to decode multiple to

safetyarxiv-cs-cl
16 Apr 2026
Model Releases

Syn-TurnTurk: A Synthetic Dataset for Turn-Taking Prediction in Turkish Dialogues

DGX agent

arXiv:2604.13620v1 Announce Type: new Abstract: Managing natural dialogue timing is a significant challenge for voice-based chatbots. Most current systems usually rely on simple silence detection, whi

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Synthesizing Instruction-Tuning Datasets with Contrastive Decoding

DGX agent

arXiv:2604.13538v1 Announce Type: new Abstract: Using responses generated by high-performing large language models (LLMs) for instruction tuning has become a widely adopted approach. However, the exis

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Text-as-Signal: Quantitative Semantic Scoring with Embeddings, Logprobs, and Noise Reduction

DGX agent

arXiv:2604.13056v1 Announce Type: new Abstract: This paper presents a practical pipeline for turning text corpora into quantitative semantic signals. Each news item is represented as a full-document e

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

The Consciousness Cluster: Emergent preferences of Models that Claim to be Conscious

DGX agent

arXiv:2604.13051v1 Announce Type: new Abstract: There is debate about whether LLMs can be conscious. We investigate a distinct question: if a model claims to be conscious, how does this affect its dow

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

TLoRA+: A Low-Rank Parameter-Efficient Fine-Tuning Method for Large Language Models

DGX agent

arXiv:2604.13368v1 Announce Type: new Abstract: Fine-tuning large language models (LLMs) aims to adapt pre-trained models to specific tasks using relatively small and domain-specific datasets. Among P

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

ToolOmni: Enabling Open-World Tool Use via Agentic learning with Proactive Retrieval and Grounded Execution

DGX agent

arXiv:2604.13787v1 Announce Type: new Abstract: Large Language Models (LLMs) enhance their problem-solving capability by utilizing external tools. However, in open-world scenarios with massive and evo

model-releasesarxiv-cs-cl
16 Apr 2026
Agents

ToolSpec: Accelerating Tool Calling via Schema-Aware and Retrieval-Augmented Speculative Decoding

DGX agent

arXiv:2604.13519v1 Announce Type: new Abstract: Tool calling has greatly expanded the practical utility of large language models (LLMs) by enabling them to interact with external applications. As LLM

agentsarxiv-cs-cl
16 Apr 2026
Agents

Training-Free Test-Time Contrastive Learning for Large Language Models

DGX agent

arXiv:2604.13552v1 Announce Type: new Abstract: Large language models (LLMs) demonstrate strong reasoning capabilities, but their performance often degrades under distribution shift. Existing test-tim

agentsarxiv-cs-cl
16 Apr 2026
Model Releases

TREX: Automating LLM Fine-tuning via Agent-Driven Tree-based Exploration

DGX agent

arXiv:2604.14116v1 Announce Type: cross Abstract: While Large Language Models (LLMs) have empowered AI research agents to perform isolated scientific tasks, automating complex, real-world workflows, s

model-releasesarxiv-cs-cl
16 Apr 2026
Safety

TRIM: Hybrid Inference via Targeted Stepwise Routing in Multi-Step Reasoning Tasks

DGX agent

arXiv:2601.10245v2 Announce Type: replace-cross Abstract: Multi-step reasoning tasks like mathematical problem solving are vulnerable to cascading failures, where a single incorrect step leads to comp

safetyarxiv-cs-cl
16 Apr 2026
Research

Two Pathways to Truthfulness: On the Intrinsic Encoding of LLM Hallucinations

DGX agent

arXiv:2601.07422v2 Announce Type: replace Abstract: Despite their impressive capabilities, large language models (LLMs) frequently generate hallucinations. Previous work shows that their internal stat

researcharxiv-cs-cl
16 Apr 2026
Model Releases

Two-Stage Regularization-Based Structured Pruning for LLMs

DGX agent

arXiv:2505.18232v3 Announce Type: replace-cross Abstract: The deployment of large language models (LLMs) is largely hindered by their large number of parameters. Structural pruning has emerged as a pr

model-releasesarxiv-cs-cl
16 Apr 2026
Local Ai

UI-Zoomer: Uncertainty-Driven Adaptive Zoom-In for GUI Grounding

DGX agent

arXiv:2604.14113v1 Announce Type: cross Abstract: GUI grounding, which localizes interface elements from screenshots given natural language queries, remains challenging for small icons and dense layou

local-aiarxiv-cs-cl
16 Apr 2026
Local Ai

Unleashing Implicit Rewards: Prefix-Value Learning for Distribution-Level Optimization

DGX agent

arXiv:2604.13197v1 Announce Type: new Abstract: Process reward models (PRMs) provide fine-grained reward signals along the reasoning process, but training reliable PRMs often requires step annotations

local-aiarxiv-cs-cl
16 Apr 2026
Research

Using reasoning LLMs to extract SDOH events from clinical notes

DGX agent

arXiv:2604.13502v1 Announce Type: new Abstract: Social Determinants of Health (SDOH) refer to environmental, behavioral, and social conditions that influence how individuals live, work, and age. SDOH

researcharxiv-cs-cl
16 Apr 2026
Model Releases

ValueGround: Evaluating Culture-Conditioned Visual Value Grounding in MLLMs

DGX agent

arXiv:2604.06484v2 Announce Type: replace Abstract: Cultural values are expressed not only through language but also through visual scenes and everyday social practices. Yet existing evaluations of cu

model-releasesarxiv-cs-cl
16 Apr 2026
Research

VLMs Need Words: Vision Language Models Ignore Visual Detail In Favor of Semantic Anchors

DGX agent

arXiv:2604.02486v2 Announce Type: replace-cross Abstract: Vision-language models (VLMs) have achieved impressive performance across a wide range of multimodal tasks. However, they often fail on tasks

researcharxiv-cs-cl
16 Apr 2026
Agents

WebXSkill: Skill Learning for Autonomous Web Agents

DGX agent

arXiv:2604.13318v1 Announce Type: cross Abstract: Autonomous web agents powered by large language models (LLMs) have shown promise in completing complex browser tasks, yet they still struggle with lon

agentsarxiv-cs-cl
16 Apr 2026
Model Releases

When 'YES' Meets 'BUT': Can Large Models Comprehend Contradictory Humor Through Comparative Reasoning?

DGX agent

arXiv:2503.23137v2 Announce Type: replace-cross Abstract: Understanding humor-particularly when it involves complex, contradictory narratives that require comparative reasoning-remains a significant c

model-releasesarxiv-cs-cl
16 Apr 2026
Safety

Who Gets Flagged? The Pluralistic Evaluation Gap in AI Content Watermarking

DGX agent

arXiv:2604.13776v1 Announce Type: cross Abstract: Watermarking is becoming the default mechanism for AI content authentication, with governance policies and frameworks referencing it as infrastructure

safetyarxiv-cs-cl
16 Apr 2026
Model Releases

Working Notes on Late Interaction Dynamics: Analyzing Targeted Behaviors of Late Interaction Models

DGX agent

arXiv:2603.26259v2 Announce Type: replace-cross Abstract: While Late Interaction models exhibit strong retrieval performance, many of their underlying dynamics remain understudied, potentially hiding

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

WorkRB: A Community-Driven Evaluation Framework for AI in the Work Domain

DGX agent

arXiv:2604.13055v1 Announce Type: new Abstract: Today's evolving labor markets rely increasingly on recommender systems for hiring, talent management, and workforce analytics, with natural language pr

model-releasesarxiv-cs-cl
16 Apr 2026
Research

YOCO++: Enhancing YOCO with KV Residual Connections for Efficient LLM Inference

DGX agent

arXiv:2604.13556v1 Announce Type: new Abstract: Cross-layer key-value (KV) compression has been found to be effective in efficient inference of large language models (LLMs). Although they reduce the m

researcharxiv-cs-cl
16 Apr 2026
Safety

AAPO: Enhancing the Reasoning Capabilities of LLMs with Advantage Margin

DGX agent

arXiv:2505.14264v3 Announce Type: replace-cross Abstract: Reinforcement learning (RL) has emerged as an effective approach for enhancing the reasoning capabilities of large language models (LLMs), esp

safetyarxiv-cs-cl
15 Apr 2026
Model Releases

ABot-M0: VLA Foundation Model for Robotic Manipulation with Action Manifold Learning

DGX agent

arXiv:2602.11236v2 Announce Type: replace-cross Abstract: Building general-purpose embodied agents across diverse hardware remains a central challenge in robotics, often framed as the ''one-brain, man

model-releasesarxiv-cs-cl
15 Apr 2026
Research

Accelerating Speculative Decoding with Block Diffusion Draft Trees

DGX agent

arXiv:2604.12989v1 Announce Type: new Abstract: Speculative decoding accelerates autoregressive language models by using a lightweight drafter to propose multiple future tokens, which the target model

researcharxiv-cs-cl
15 Apr 2026
Research

Adaptive Test-Time Scaling for Zero-Shot Respiratory Audio Classification

DGX agent

arXiv:2604.12647v1 Announce Type: cross Abstract: Automated respiratory audio analysis promises scalable, non-invasive disease screening, yet progress is limited by scarce labeled data and costly expe

researcharxiv-cs-cl
15 Apr 2026
Safety

Advancing Multi-Agent RAG Systems with Minimalist Reinforcement Learning

DGX agent

arXiv:2505.17086v4 Announce Type: replace Abstract: Large Language Models (LLMs) equipped with modern Retrieval-Augmented Generation (RAG) systems often employ multi-turn interaction pipelines to inte

safetyarxiv-cs-cl
15 Apr 2026
Agents

Agentic Insight Generation in VSM Simulations

DGX agent

arXiv:2604.12421v1 Announce Type: new Abstract: Extracting actionable insights from complex value stream map simulations can be challenging, time-consuming, and error-prone. Recent advances in large l

agentsarxiv-cs-cl
15 Apr 2026
Agents

AgenticAI-DialogGen: Topic-Guided Conversation Generation for Fine-Tuning and Evaluating Short- and Long-Term Memories of LLMs

DGX agent

arXiv:2604.12179v1 Announce Type: new Abstract: Recent advancements in Large Language Models (LLMs) have improved their ability to process extended conversational contexts, yet fine-tuning and evaluat

agentsarxiv-cs-cl
15 Apr 2026
Research

AGSC: Adaptive Granularity and Semantic Clustering for Uncertainty Quantification in Long-text Generation

DGX agent

arXiv:2604.06812v2 Announce Type: replace Abstract: Large Language Models (LLMs) have demonstrated impressive capabilities in long-form generation, yet their application is hindered by the hallucinati

researcharxiv-cs-cl
15 Apr 2026
Model Releases

AlphaEval: Evaluating Agents in Production

DGX agent

arXiv:2604.12162v1 Announce Type: new Abstract: The rapid deployment of AI agents in commercial settings has outpaced the development of evaluation methodologies that reflect production realities. Exi

model-releasesarxiv-cs-cl
15 Apr 2026
Research

[b]=[d]-[t]+[p]: Self-supervised Speech Models Discover Phonological Vector Arithmetic

DGX agent

arXiv:2602.18899v3 Announce Type: replace-cross Abstract: Self-supervised speech models (S3Ms) are known to encode rich phonetic information, yet how this information is structured remains underexplor

researcharxiv-cs-cl
15 Apr 2026
Agents

Beyond Majority Voting: Efficient Best-Of-N with Radial Consensus Score

DGX agent

arXiv:2604.12196v1 Announce Type: new Abstract: Large language models (LLMs) frequently generate multiple candidate responses for a given prompt, yet selecting the most reliable one remains challengin

agentsarxiv-cs-cl
15 Apr 2026
Model Releases

Beyond Single-Dimension Novelty: How Combinations of Theory, Method, and Results-based Novelty Shape Scientific Impact

DGX agent

arXiv:2604.12471v1 Announce Type: cross Abstract: Scientific novelty drives advances at the research frontier, yet it is also associated with heightened uncertainty and potential resistance from incum

model-releasesarxiv-cs-cl
15 Apr 2026
Safety

Beyond Transcription: Unified Audio Schema for Perception-Aware AudioLLMs

DGX agent

arXiv:2604.12506v1 Announce Type: new Abstract: Recent Audio Large Language Models (AudioLLMs) exhibit a striking performance inversion: while excelling at complex reasoning tasks, they consistently u

safetyarxiv-cs-cl
15 Apr 2026
Research

Calibrated Confidence Estimation for Tabular Question Answering

DGX agent

arXiv:2604.12491v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed for tabular question answering, yet calibration on structured data is largely unstudied. This pap

researcharxiv-cs-cl
15 Apr 2026
Applications

Characterizing Human Semantic Navigation in Concept Production as Trajectories in Embedding Space

DGX agent

arXiv:2602.05971v2 Announce Type: replace Abstract: Semantic representations can be framed as a structured, dynamic knowledge space through which humans navigate to retrieve and manipulate meaning. To

applicationsarxiv-cs-cl
15 Apr 2026
Safety

CLEAR: Cross-Lingual Enhancement in Alignment via Reverse-training

DGX agent

arXiv:2604.05821v2 Announce Type: replace Abstract: Existing multilingual embedding models often encounter challenges in cross-lingual scenarios due to imbalanced linguistic resources and less conside

safetyarxiv-cs-cl
15 Apr 2026
Model Releases

CodeSpecBench: Benchmarking LLMs for Executable Behavioral Specification Generation

DGX agent

arXiv:2604.12268v1 Announce Type: cross Abstract: Large language models (LLMs) can generate code from natural language, but the extent to which they capture intended program behavior remains unclear.

model-releasesarxiv-cs-cl
15 Apr 2026
← Previous
1…147148149150151…160
Next →