AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,759 results
22 May 2026

SpaceDG: Benchmarking Spatial Intelligence under Visual Degradation

Model ReleasesDGX agent

arXiv:2605.22536v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have made rapid progress in spatial intelligence, yet existing spatial reasoning benchmarks largely assume pr

SpecHop: Continuous Speculation for Accelerating Multi-Hop Retrieval Agents

AgentsDGX agent

arXiv:2605.21965v1 Announce Type: new Abstract: Large language models increasingly use external tools such as web search and document retrieval to solve information-intensive tasks. However, multi-hop

STRUCTSENSE: A Task-Agnostic Agentic Framework for Structured Information Extraction with Human-In-The-Loop Evaluation and Benchmarking

AgentsDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2507.03674v3 Announce Type: replace Abstract: Extracting structured information from scientific literature is critical for accelerating discovery, yet Large Language Models (LLMs) often struggle

Structural Anchor Pruning: Training-Free Multi-Vector Compression for Visual Document Retrieval

Model ReleasesDGX agent

arXiv:2601.20107v2 Announce Type: replace-cross Abstract: Recent Vision-Language Models (e.g., ColPali) enable fine-grained Visual Document Retrieval (VDR) but incur prohibitive multi-vector index sto

Structure Retention in Embedding Spaces as a Predictor of Benchmark Performance

Model ReleasesDGX agent

arXiv:2605.22202v1 Announce Type: new Abstract: In this paper, we show that high-performing embedding models organize their embedding spaces in a consistent way. We evaluate 25 contemporary embedding

Structured-Sparse Attention for Entity Tracking with Subquadratic Sequence Complexity

Local AiDGX agent

arXiv:2605.22476v1 Announce Type: cross Abstract: Entity tracking requires maintaining and updating latent states for entities and attributes over long sequences. Recent task-specific attention operat

Sub-exponential Growth Dynamics in Complex Systems: A Piecewise Power-Law Model for the Diffusion of New Words and Names

Model ReleasesDGX agent

arXiv:2511.04106v5 Announce Type: replace-cross Abstract: The diffusion of ideas and language in society has conventionally been described by S-shaped models, such as the logistic curve. However, the

Survive or Collapse: The Asymmetric Roles of Data Gating and Reward Grounding in Self-Play RL

Model ReleasesDGX agent

arXiv:2605.22217v1 Announce Type: cross Abstract: Self-play reinforcement learning trains language models on their own generated tasks, co-evolving a proposer and solver without human labels. Recent s

SynAE: A Framework for Measuring the Quality of Synthetic Data for Tool-Calling Agent Evaluations

AgentsDGX agent

arXiv:2605.22564v1 Announce Type: new Abstract: Today, tool-calling agents are commonly evaluated or tested on static datasets of execution traces, including input commands, agent responses, and assoc

Teaching Language Models to Forecast Research Success Through Comparative Idea Evaluation

Model ReleasesDGX agent

arXiv:2605.21491v1 Announce Type: cross Abstract: As language models accelerate scientific research by automating hypothesis generation and implementation, a new bottleneck emerges: evaluating and fil

The Double Dilemma in Multi-Task Radiology Report Generation: A Gradient Dynamics Analysis and Solution

SafetyDGX agent

arXiv:2605.22635v1 Announce Type: cross Abstract: While multi-task learning based automatic radiology report generation (RRG) is widely adopted to ensure clinical consistency, most focus on architectu

Token-Level LLM Collaboration via FusionRoute

Model ReleasesDGX agent

arXiv:2601.05106v4 Announce Type: replace-cross Abstract: Large language models (LLMs) exhibit strengths across diverse domains. However, achieving strong performance across these domains with a singl

Token-weighted Direct Preference Optimization with Attention

ResearchDGX agent

arXiv:2605.21883v1 Announce Type: new Abstract: Direct Preference Optimization (DPO) aligns Large Language Models with human preferences without the need for a separate reward model. However, DPO trea

Tokenisation via Convex Relaxations

ResearchDGX agent

arXiv:2605.22821v1 Announce Type: new Abstract: Tokenisation is an integral part of the current NLP pipeline. Current tokenisation algorithms such as BPE and Unigram are greedy algorithms -- they make

Tokenization with Split Trees

Model ReleasesDGX agent

arXiv:2605.22705v1 Announce Type: new Abstract: We introduce Tokenization with Split Trees (ToaST), a subword tokenization method that directly optimizes compression under a new recursive inference pr

Training-Trajectory-Aware Token Selection

Model ReleasesDGX agent

arXiv:2601.10348v2 Announce Type: replace Abstract: Efficient distillation is a key pathway for converting expensive reasoning capability into deployable efficiency, yet in the frontier regime where t

TransitLM: A Large-Scale Dataset and Benchmark for Map-Free Transit Route Generation

Model ReleasesDGX agent

arXiv:2605.22355v1 Announce Type: new Abstract: Public transit route planning traditionally depends on structured map infrastructure and complex routing engines, and no existing dataset supports train

Two is better than one: A Collapse-free Multi-Reward RLIF Training Framework

Model ReleasesDGX agent

arXiv:2605.22620v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has substantially improved the reasoning ability of LLMs, but often depends on external supervis

Understanding Data Temporality Impact on Large Language Models Pre-training

Model ReleasesDGX agent

arXiv:2605.22769v1 Announce Type: new Abstract: Large language models (LLMs) are typically trained on shuffled corpora, yielding models whose knowledge is frozen at train time and whose temporal groun

Unified Data Selection for LLM Reasoning

TutorialsDGX agent

arXiv:2605.22389v1 Announce Type: new Abstract: Effectively training Large Language Models (LLMs) for complex, long-CoT reasoning is often bottlenecked by the need for massive high-quality reasoning d

Unifying Masked Diffusion Models with Various Generation Orders and Beyond

SafetyDGX agent

arXiv:2602.02112v2 Announce Type: replace-cross Abstract: Masked diffusion models (MDMs) are a potential alternative to autoregressive models (ARMs) for language generation, but generation quality dep

UniSD: Towards a Unified Self-Distillation Framework for Large Language Models

SafetyDGX agent

arXiv:2605.06597v2 Announce Type: replace Abstract: Self-distillation (SD) offers a promising path for adapting large language models (LLMs) without relying on stronger external teachers. However, SD

Value-Gradient Hypothesis of RL for LLMs

SafetyDGX agent

arXiv:2605.21654v1 Announce Type: cross Abstract: Reinforcement learning substantially improves pretrained language models, but it remains understudied why critic-free methods such as PPO and GRPO wor

Vector Policy Optimization: Training for Diversity Improves Test-Time Search

SafetyDGX agent

arXiv:2605.22817v1 Announce Type: cross Abstract: Language models must now generalize out of the box to novel environments and work inside inference-scaling search procedures, such as AlphaEvolve, tha

When Cases Get Rare: A Retrieval Benchmark for Off-Guideline Clinical Question Answering

Model ReleasesDGX agent

arXiv:2605.21807v1 Announce Type: new Abstract: Across medical specialties, clinical practice is anchored in evidence-based guidelines that codify best studied diagnostic and treatment pathways. These

When Shared Knowledge Hurts: Spectral Over-Accumulation in Model Merging

ResearchDGX agent

arXiv:2602.05536v2 Announce Type: replace-cross Abstract: Model merging combines multiple fine-tuned models into a single model by adding their weight updates, providing a lightweight alternative to r

Whose Voice Counts? Mapping Stakeholder Perspectives on AI Through Public Submissions to the U.S. Government

ResearchDGX agent

arXiv:2605.22650v1 Announce Type: new Abstract: As artificial intelligence (AI) systems become more common in our daily lives, it is important to understand how different stakeholders comprehend and e

Why Semantic Entropy Fails: Geometry-Aware and Calibrated Uncertainty for Policy Optimization

SafetyDGX agent

arXiv:2605.21801v1 Announce Type: cross Abstract: Post-training has become central to improving reasoning and alignment in large language models, where critic-free models enable scalable learning from

'Would You Want an AI Tutor?' Understanding Stakeholder Perceptions of LLM-based Systems in the Classroom

SafetyDGX agent

arXiv:2503.02885v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have gained traction in educational settings, often framed as virtual tutors or teaching assistants. Following ea

X-Token: Projection-Guided Cross-Tokenizer Knowledge Distillation

Model ReleasesDGX agent

arXiv:2605.21699v1 Announce Type: cross Abstract: Cross-tokenizer knowledge distillation allows a student model to learn from teachers with incompatible vocabularies. Prior work operates on hidden sta

21 May 2026

A Systematic Comparison between Extractive Self-Explanations and Human Rationales in Text Classification

SafetyDGX agent

arXiv:2410.03296v4 Announce Type: replace Abstract: Instruction-tuned LLMs are able to provide extit{an} explanation about their output to users by generating self-explanations, without requiring the

ACL-Verbatim: hallucination-free question answering for research

Model ReleasesDGX agent

arXiv:2605.21102v1 Announce Type: new Abstract: Academic researchers need efficient and reliable methods for collecting high-quality information from trusted sources, but modern tools for AI-assisted

AFD-INSTRUCTION: A Comprehensive Antibody Instruction Dataset with Functional Annotations for LLM-Based Understanding and Design

SafetyDGX agent

arXiv:2602.04916v3 Announce Type: replace-cross Abstract: Large language models (LLMs) have significantly advanced protein representation learning. However, their capacity to interpret and design anti

AgentAtlas: Beyond Outcome Leaderboards for LLM Agents

Model ReleasesDGX agent

arXiv:2605.20530v1 Announce Type: cross Abstract: Large language model agents now act on codebases, browsers, operating systems, calendars, files, and tool ecosystems, but the benchmarks used to evalu

AI-Assisted Scientific Assessment: A Case Study on Climate Change

Model ReleasesDGX agent

arXiv:2602.09723v2 Announce Type: replace Abstract: The emerging paradigm of AI co-scientists focuses on tasks characterized by repeatable verification, where agents explore search spaces in 'guess an

AI-Augmented Surveys: Leveraging Large Language Models and Surveys for Opinion Prediction

ResearchDGX agent

arXiv:2305.09620v4 Announce Type: replace Abstract: Nationally representative surveys track public opinion, yet they ask only a limited set of questions each year, limiting its potential to capture hi

AiraXiv: An AI-Driven Open-Access Platform for Human and AI Scientists

AgentsDGX agent

arXiv:2605.21481v1 Announce Type: cross Abstract: Recent advances in artificial intelligence (AI) have accelerated the growth of both human-authored and AI-generated research outputs, placing increasi

Anatomy of Agentic Memory: Taxonomy and Empirical Analysis of Evaluation and System Limitations

Model ReleasesDGX agent

arXiv:2602.19320v2 Announce Type: replace Abstract: Agentic memory systems enable large language model (LLM) agents to maintain state across long interactions, supporting long-horizon reasoning and pe

Anti-establishment sentiment on TikTok: Implications for understanding influence(rs) and expertise on social media

ApplicationsDGX agent

arXiv:2508.16453v2 Announce Type: replace-cross Abstract: Distrust of public serving institutions and anti-establishment views are on the rise (especially in the U.S.). As people turn to social media

APM: Evaluating Style Personalization in LLMs with Arbitrary Preference Mappings

Model ReleasesDGX agent

arXiv:2605.21063v1 Announce Type: new Abstract: Typical LLM responses tend to follow a default style, even though users often have distinct preferences regarding tone, verbosity, and formality that th

ArPoMeme: An Annotated Arabic Multimodal Dataset for Political Ideology and Polarization

ResearchDGX agent

arXiv:2605.20967v1 Announce Type: new Abstract: Memes have become a prominent medium of political communication in the Arab world, reflecting how humor, imagery, and text interact to express ideologic

Assessing socio-economic climate impacts from text data

ResearchDGX agent

arXiv:2605.20793v1 Announce Type: new Abstract: Recent advances in natural language processing (NLP) and large language models (LLMs) have enabled the systematic use of large-scale textual data from n

Auto-Dreamer: Learning Offline Memory Consolidation for Language Agents

AgentsDGX agent

arXiv:2605.20616v1 Announce Type: new Abstract: Language agents increasingly operate over streams of related tasks, yet existing memory systems struggle to convert accumulated experience into reusable

Automated ICD Classification of Psychiatric Diagnoses: From Classical NLP to Large Language Models

Model ReleasesDGX agent

arXiv:2605.21154v1 Announce Type: new Abstract: Mental health has become a global priority, leading to a massive administrative burden in the coding of clinical diagnoses. This study proposes the auto

Automatically Learning Construction Injury Precursors from Text

SafetyDGX agent

arXiv:1907.11769v4 Announce Type: replace Abstract: In light of the increasing availability of digitally recorded safety reports in the construction industry, it is important to develop methods to exp

AVSD: Adaptive-View Self-Distillation by Balancing Consensus and Teacher-Specific Privileged Signals

SafetyDGX agent

arXiv:2605.20643v1 Announce Type: cross Abstract: Self-distillation enables language models to learn on-policy from their own trajectories by using the same model as both student and teacher, with the

Bayesian Preference Learning for Test-Time Steerable Reward Models

SafetyDGX agent

arXiv:2602.08819v2 Announce Type: replace-cross Abstract: Reward models are central to aligning language models with human preferences via reinforcement learning (RL). As RL is increasingly applied to

Beyond Semantic Similarity: A Two-Phase Non-Parametric Retrieval Workflow for Corporate Credit Underwriting

ApplicationsDGX agent

arXiv:2605.20684v1 Announce Type: new Abstract: Corporate credit underwriting requires analysts to extract actionable evidence from long, heterogeneous financial documents spanning hundreds of pages a

Beyond Text-to-SQL: An Agentic LLM System for Governed Enterprise Analytics APIs

SafetyDGX agent

arXiv:2605.21027v1 Announce Type: new Abstract: Enterprise analytics aims to make organizational data accessible for decision-making, yet non-technical users still face barriers when using traditional

Beyond Words: Multimodal LLM Knows When to Speak

AgentsDGX agent

arXiv:2505.14654v2 Announce Type: replace-cross Abstract: Chatbots via large language models (LLMs) generate fluent responses but often struggle with when to speak, especially for brief, timely listen

Bridging Language Models and Financial Analysis

ApplicationsDGX agent

arXiv:2503.22693v2 Announce Type: replace-cross Abstract: The rapid advancements in Large Language Models (LLMs) have unlocked transformative possibilities in natural language processing, particularly

Building a Custom Taxonomy of AI Skills and Tasks from the Ground Up with Job Postings

TutorialsDGX agent

arXiv:2605.21029v1 Announce Type: new Abstract: Utilizing LLMs for automated taxonomy construction presents a clear opportunity for the comprehensive, yet efficient mapping of potentially complex doma

Building Arabic NLP from the Ground Up: Twenty Years of Lessons, Failures, and Open Problems

ResearchDGX agent

arXiv:2605.20786v1 Announce Type: new Abstract: This paper reflects on twenty years of building NLP resources and research infrastructure for Arabic, a language spoken by hundreds of millions yet hist

Calibration vs Decision Making: Revisiting the Reliability Paradox in Unlearned Language Models

Model ReleasesDGX agent

arXiv:2605.20915v1 Announce Type: new Abstract: Machine unlearning aims to remove the influence of specific training data from a model while preserving reliable behavior on the remaining data, making

Causal Path Alignment: Anchoring the Optimization Trajectory for Controllable In-Parameter Knowledge Editing

Model ReleasesDGX agent

arXiv:2506.04042v2 Announce Type: replace Abstract: Knowledge editing is pivotal for efficiently updating the parametric memory of Large Language Models (LLMs), enabling them to function as evolving a

Chronicle: A Multimodal Foundation Model for Joint Language and Time Series Understanding

Model ReleasesDGX agent

arXiv:2605.20268v1 Announce Type: cross Abstract: Real-world time series come with text: metadata, descriptions, news, reports. Yet time series foundation models process numerical sequences in isolati

ChunkFT: Byte-Streamed Optimization for Memory-Efficient Full Fine-Tuning

Model ReleasesDGX agent

arXiv:2605.21177v1 Announce Type: cross Abstract: This work presents extsc{ChunkFT}, a memory-efficient fine-tuning framework that reformulates full-parameter fine-tuning around a dynamically activate

Collocational bootstrapping: A hypothesis about the learning of subject-verb agreement in humans and neural networks

TutorialsDGX agent

arXiv:2605.20529v1 Announce Type: new Abstract: In what ways might statistical signals in linguistic input assist with the acquisition of syntax? Here we hypothesize a mechanism called collocational b

CP-MoE: Consistency-Preserving Mixture-of-Experts for Continual Learning

Model ReleasesDGX agent

arXiv:2605.20247v1 Announce Type: cross Abstract: Catastrophic forgetting remains a major obstacle to continual learning in large language models (LLMs) and vision--language models (VLMs). Although Mi

Cross-lingual robustness of LLM-brain alignment and its computational roots

SafetyDGX agent

arXiv:2605.21049v1 Announce Type: new Abstract: Large language models (LLMs) reliably predict neural activity during language comprehension and transformer depth has been interpreted as mirroring hier

← Previous
1…6667686970…130
Next →