AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
29 May 2026

Casual as an Anchor: Resolving Supervision Misalignment in Formality Transfer Dataset

Model ReleasesDGX agent

arXiv:2605.29365v1 Announce Type: new Abstract: Formality transfer is commonly framed as a symmetric bidirectional task between informal and formal registers. We argue that this framing conceals a sup

Catalyst-Agent: Autonomous heterogeneous catalyst screening with an LLM Agent

AgentsDGX agent

arXiv:2603.01311v2 Announce Type: replace Abstract: The discovery of novel catalysts tailored for particular applications is a major challenge for the twenty-first century. Traditional methods for thi

Causal Interventions on Continuous Variables: A Case Study on Verb Bias in Steering Vectors for In-Context Learning

SafetyDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.29971v1 Announce Type: new Abstract: Causal interventions in language model representations have largely targeted discrete features, like grammatical number. However, language models must a

CCS: Clinical Consensus Selection for Radiology Report Generation

ResearchDGX agent

arXiv:2605.30131v1 Announce Type: new Abstract: Radiology report generation (RRG) is commonly formulated as a single-path generation task, where a multimodal large language model (MLLM) produces one d

Classification of non-analyzable word types in web documents to implement an effective Korean e-learning system

Local AiDGX agent

arXiv:2605.29638v1 Announce Type: new Abstract: E-learning systems should deliver contents that reflect various phenomena of the language as it is used. In addition to formal Korean, e-learning system

Cognitive Loop of Thought: Reversible Hierarchical Markov Chain for Efficient Mathematical Reasoning

ResearchDGX agent

arXiv:2604.06805v2 Announce Type: replace Abstract: Multi-step Chain-of-Thought (CoT) has significantly advanced the mathematical reasoning capabilities of LLMs by leveraging explicit reasoning steps.

CommunityFact: A Dynamic, Multilingual, Multi-domain Benchmark for Misinformation Detection in the Wild

Model ReleasesDGX agent

arXiv:2605.30241v1 Announce Type: new Abstract: Misinformation verification increasingly occurs in public, fast-moving, and multilingual online settings, where static benchmarks provide an incomplete

Comparative Evaluation of Machine Translation Systems on Images with Text

Model ReleasesDGX agent

arXiv:2605.29476v1 Announce Type: new Abstract: This work presents a comparative evaluation of machine translation systems applied to images containing textual information, a task that lies at the int

COMPOSE: Composing Future Theorems from Citations and Formal Structure

Model ReleasesDGX agent

arXiv:2605.30333v1 Announce Type: new Abstract: A plausible future mathematical claim must satisfy two constraints: it should follow the direction of prior work and respect the formal dependencies tha

CONCAT: Consensus- and Confidence-Driven Ad Hoc Teaming for Efficient LLM-Based Multi-Agent Systems

AgentsDGX agent

arXiv:2605.29612v1 Announce Type: cross Abstract: Although large language model (LLM) based multi-agent systems (MAS) show their capability to solve complex tasks and achieve higher performance over s

Converted, Not Equivalent: Benchmarking Codebase Conversion via Observational Equivalence

Model ReleasesDGX agent

arXiv:2605.29054v1 Announce Type: cross Abstract: Coding agents increasingly act as codebase-scale collaborators that can assist with codebase conversion, but this progress has exposed a critical weak

CorPipe at CRAC 2026: Empty Nodes and Cross-Lingual Transfer in Multilingual Coreference Resolution

ResearchDGX agent

arXiv:2605.30133v1 Announce Type: new Abstract: We introduce CorPipe 26, our winning submission to the CRAC 2026 Shared Task on Multilingual Coreference Resolution. The fifth edition of this shared ta

CriticalKV: Optimizing KV Cache Eviction from an Output Perturbation Perspective

Model ReleasesDGX agent

arXiv:2502.03805v2 Announce Type: replace Abstract: Large language models have revolutionized natural language processing but face significant challenges of high storage and runtime costs, due to the

Demystifying Scientific Problem-Solving in LLMs by Probing Knowledge and Reasoning

Model ReleasesDGX agent

arXiv:2508.19202v3 Announce Type: replace Abstract: Scientific problem solving poses unique challenges for LLMs, requiring both deep domain knowledge and the ability to apply such knowledge through co

DFlash: Block Diffusion for Flash Speculative Decoding

HardwareDGX agent

arXiv:2602.06036v2 Announce Type: replace Abstract: Autoregressive large language models (LLMs) deliver strong performance but require inherently sequential decoding, leading to high inference latency

Dial HEALTHDIAL for Advice: A Multilingual and Multi-Parallel Spoken Dialogue Dataset for Knowledge-Grounded Information Seeking

Model ReleasesDGX agent

arXiv:2605.30107v1 Announce Type: new Abstract: Creating spoken dialogue datasets is methodologically challenging, and these challenges are amplified when the goal is to build multilingual, multi-para

DiffSpot: Can VLMs Spot Fine-Grained Visual Differences in Web Interfaces?

Model ReleasesDGX agent

arXiv:2605.29615v1 Announce Type: cross Abstract: Vision-language models (VLMs) have made strong progress on high-level image-text alignment, yet their ability to perceive subtle visual differences re

DirectorBench: Diagnosing Long-Form Video Generation with Personalized Multi-Agent Evaluation

Model ReleasesDGX agent

arXiv:2605.30090v1 Announce Type: new Abstract: Long-form video generation is rapidly moving from short, single-scene synthesis toward minute-long, multi-shot creation with narrative structure, cinema

DLT-Corpus: A Large-Scale Text Collection for the Distributed Ledger Technology Domain

ResearchDGX agent

arXiv:2602.22045v2 Announce Type: replace Abstract: We introduce DLT-Corpus, the largest domain-specific text collection for Distributed Ledger Technology (DLT) research to date: 2.98 billion tokens f

Do not be greedy, Think Twice: Sampling and Selection for Document-level Information Extraction

ResearchDGX agent

arXiv:2601.18395v2 Announce Type: replace Abstract: Document-level Information Extraction (DocIE) aims to produce an output template with the entities, relations, and events of interest occurring in t

Domino: Decoupling Causal Modeling from Autoregressive Drafting in Speculative Decoding

ResearchDGX agent

arXiv:2605.29707v1 Announce Type: new Abstract: Speculative decoding accelerates LLM inference by drafting multiple tokens and verifying them in parallel with the target model. However, its practical

Draft-OPD: On-Policy Distillation for Speculative Draft Models

SafetyDGX agent

arXiv:2605.29343v1 Announce Type: new Abstract: Speculative decoding accelerates large language model inference by pairing a target model with a lightweight draft model whose proposed tokens are verif

DynaGraph: Lightweight Multi-Model Interaction Framework via Dynamic Topological Reconfiguration

Local AiDGX agent

arXiv:2605.29511v1 Announce Type: cross Abstract: Tackling complex reasoning tasks typically relies on massive monolithic LLMs, which suffer from severe computational redundancy. While task decomposit

DySem: Uncovering Dynamic Semantic Components via Multilingual Consensus for Calculating Semantic Textual Similarity

Model ReleasesDGX agent

arXiv:2605.29751v1 Announce Type: new Abstract: Calculating semantic textual similarity is a foundational task in natural language processing. Current large language models (LLMs) based methods typica

Early Detection of Misinformation for Infodemic Management: A Domain Adaptation Approach

TutorialsDGX agent

arXiv:2406.10238v2 Announce Type: replace Abstract: An infodemic refers to an enormous amount of true information and misinformation disseminated during a disease outbreak. Detecting misinformation at

Efficient Training-Free Multi-Token Prediction via Embedding-Space Probing

ResearchDGX agent

arXiv:2603.17942v2 Announce Type: replace Abstract: Large Language Models (LLMs) possess latent multi-token prediction (MTP) abilities despite being trained only for next-token generation. We introduc

Enhancing Factuality through Consensus and Consistency in Summarization Using Minimum Bayes Risk Decoding

ResearchDGX agent

arXiv:2605.29336v1 Announce Type: new Abstract: Improving the quality of model-generated summaries, especially factuality, the accuracy of a summary with respect to its source content, remains a chall

Error as a Lens: Probing LLM Reasoning through Synthetic Misconception Generation

AgentsDGX agent

arXiv:2605.29007v1 Announce Type: new Abstract: Personalized tutoring, teacher training, and education research need access to targeted synthetic misconceptions, but privacy and IRB constraints make l

EVADE: LLM-Based Explanation Generation and Validation for Error Detection in NLI

ResearchDGX agent

arXiv:2511.08949v2 Announce Type: replace Abstract: High-quality datasets are critical for training and evaluating reliable NLP models. In tasks like natural language inference (NLI), human label vari

Evaluating Cross-lingual Knowledge Consistency in Code-Mixed vis-a-vis Indian Languages using IndicKLAR

Model ReleasesDGX agent

arXiv:2605.29637v1 Announce Type: new Abstract: Large language models recall knowledge reliably in English but often fail on the same query posed in a lower-resourced language -- a crosslingual consis

EvoRubric: Self-Evolving Rubric-Driven RL for Open-Ended Generation

SafetyDGX agent

arXiv:2605.29847v1 Announce Type: new Abstract: Reinforcement Learning (RL) has significantly advanced Large Language Models (LLMs) in verifiable domains, but aligning models for open-ended generation

ExCAM: Explainable Cultural Awareness Metrics

Model ReleasesDGX agent

arXiv:2605.29897v1 Announce Type: new Abstract: Evaluating the cultural awareness of large language models is crucial to ensure the fairness of generated text and the generalizability of applications

FinGuard: Detecting Financial Regulatory Non-Compliance in LLM Interactions

Model ReleasesDGX agent

arXiv:2605.29427v1 Announce Type: new Abstract: As large language models (LLMs) are increasingly deployed in financial services, a single non-compliant interaction can expose institutions to regulator

FoRA: Fisher-orthogonal Rank Adaptation for Parameter-Efficient Fine-Tuning

Model ReleasesDGX agent

arXiv:2605.29317v1 Announce Type: new Abstract: Parameter-efficient fine-tuning(PEFT) has largely focused on LoRA and its accuracy-oriented variants, leaving the original goal of reducing trainable pa

From Blind Guess to Informed Judgment: Teaching LLMs to Evaluate Materials by Building Knowledge-Augmented Preference Signals

AgentsDGX agent

arXiv:2605.29555v1 Announce Type: new Abstract: As candidate generation and high-throughput experimentation advance, the primary bottleneck in materials discovery is shifting from property prediction

From Context Shift to Stylistic Collapse: Why Training Objectives Matter More Than Scale

SafetyDGX agent

arXiv:2605.28826v1 Announce Type: new Abstract: In modern LLMs, linguistic features function not as stylistic artifacts but as probes of probability mass, allocated under training alignment objectives

From Data to Insights: Exploring Program-of-Thoughts Prompting for Chart Summarization

ResearchDGX agent

arXiv:2605.28874v1 Announce Type: new Abstract: Charts play a critical role in conveying numerical data insights through structured visual representations. However, semantic visual understanding and n

GAPD: Gold-Action Policy Distillation for Agentic Reinforcement Learning in Knowledge Base Question Answering

SafetyDGX agent

arXiv:2605.29584v1 Announce Type: new Abstract: Reinforcement learning (RL) is a natural fit for agentic knowledge base question answering (KBQA), where a model must issue executable actions, observe

GRASP: Plan-Guided Graph Retrieval with Adaptive Fusion and Reranking on Semi-Structured Knowledge Bases

ResearchDGX agent

arXiv:2605.30237v1 Announce Type: cross Abstract: Semi-structured knowledge bases (SKBs) embed textual documents in a typed graph of entities and relations, and underpin applications such as product s

GRUFF: LLM Pronoun Fidelity, Reasoning, and Biases in German

SafetyDGX agent

arXiv:2605.30214v1 Announce Type: new Abstract: Third-person singular pronouns have long been used to study stereotypical biases in language models and to test their abilities to reason about referenc

HaluNet: Learning Hallucination Risk from Internal Signals in LLM Question Answering

ResearchDGX agent

arXiv:2512.24562v2 Announce Type: replace Abstract: Large language models (LLMs) achieve strong question answering (QA) performance but can produce fluent answers unsupported by available evidence. Ex

HEART-Bench: Do LLM Agents Exhibit Human-like Psychology?

Model ReleasesDGX agent

arXiv:2605.30058v1 Announce Type: new Abstract: While LLM agents have demonstrated remarkable task-oriented abilities such as planning, reasoning, and action, few works have treated them as complete h

How Far Ahead Do LLMs Plan? Uncovering the Latent Horizon in Chain-of-Thought Reasoning

Model ReleasesDGX agent

arXiv:2602.02103v2 Announce Type: replace-cross Abstract: Chain-of-thought (CoT) reasoning has become a central mechanism for eliciting multi-step reasoning in Large Language Models (LLMs). Yet recent

How's it going? Reinforcement learning in language models recruits a functional welfare axis

SafetyDGX agent

arXiv:2605.30232v1 Announce Type: cross Abstract: How does reinforcement learning shape a language model's internal representations? We present evidence that RL recruits a pre-existing representation

HTAM: Hierarchical Transition-Attended Memory for Operator Optimization

Local AiDGX agent

arXiv:2605.29734v1 Announce Type: new Abstract: High-performance GPU kernels are essential for efficient LLM deployment, yet optimizing them remains expertise-intensive. Recent LLM-based code generati

Implicit Identity Technologies for LLMs: Fingerprinting and Watermarking across Datasets, Models, and Generated Content

ResearchDGX agent

arXiv:2605.29245v1 Announce Type: cross Abstract: This paper presents a survey and taxonomy of LLM fingerprinting and watermarking for identity, ownership verification, provenance, and generated-conte

Interactive In-Meeting Speaker Correction with Human Feedback

ResearchDGX agent

arXiv:2509.18377v2 Announce Type: replace Abstract: Most automatic speech processing systems operate in ``open loop'' mode without user feedback about who said what, yet human-in-the-loop workflows ca

Knowing What to Solve Before How: Preplan Empowered LLM Mathematical Reasoning

TutorialsDGX agent

arXiv:2605.30245v1 Announce Type: new Abstract: Current plan-based reasoning methods improve large language models (LLMs) by inserting a planning stage before execution, giving rise to the question ri

Kronecker Embeddings: Byte-Level Structured Token Representations for Parameter-Efficient Language Models

Model ReleasesDGX agent

arXiv:2605.29459v1 Announce Type: new Abstract: Large language models route every input through a learned embedding table of shape |V| x d_model, consuming hundreds of millions to billions of trainabl

Large language models reorganize representational geometry during in-context learning

Model ReleasesDGX agent

arXiv:2605.28854v1 Announce Type: new Abstract: Large language models (LLMs) exhibit remarkable flexibility: they can adapt to novel tasks from in-context examples without any parameter updates, a cap

Latent Performance Profiling of Large Language Models

Model ReleasesDGX agent

arXiv:2605.30018v1 Announce Type: new Abstract: Large language models (LLMs) frequently achieve impressive scores on standardized benchmarks, yet accuracy alone offers a limited view of their capabili

Learnable Assessment Skills for LLM-based Automated Scoring: Rubric Construction via Iterative Optimization

TutorialsDGX agent

arXiv:2605.29274v1 Announce Type: new Abstract: LLM-based automated scoring approaches near-human performance, but scaling to new tasks remains bottlenecked by the per-item human configuration of upst

Learning Design Skills as Memory Policies for Agentic Photonic Inverse Design

Model ReleasesDGX agent

arXiv:2605.29421v1 Announce Type: new Abstract: Photonic crystal fiber (PCF) inverse design remains challenging because candidate geometries must satisfy coupled optical targets under expensive electr

Leveraging Routing Dynamics in Mixture-of-Experts Models for Efficient Language Adaptation

Model ReleasesDGX agent

arXiv:2605.29714v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) models are widely used to scale language models, yet their expert routing behavior and adaptation in a multilingual setting rem

Lexical categories of stem-forming roots in Mapudungun verb forms

ResearchDGX agent

arXiv:2502.07623v4 Announce Type: replace Abstract: After developing a computational system for morphological analysis of the Mapuche language, and evaluating it with texts from various authors and st

Lightweight Multimodal LLM-Enabled Cost-Effective Defect Grading of Power Transmission Equipment

ResearchDGX agent

arXiv:2605.28822v1 Announce Type: new Abstract: Defect grading of power transmission equipment (DGPTE) is crucial to the stability of electric energy transmission. Although existing machine learning m

LiteCoder-Terminal: Scaling Long-Horizon Terminal Environments for Learning Language Agents

Model ReleasesDGX agent

arXiv:2605.29559v1 Announce Type: new Abstract: Mastering terminal environments requires language agents capable of multi-step planning, feedback-grounded execution, and dynamic state adaptation. Howe

LLMBridge: An LLM Pipeline for End-to-end Referential Bridging Resolution in English

ResearchDGX agent

arXiv:2605.29048v1 Announce Type: new Abstract: In this paper, we introduce LLMBridge, a new LLM based system for the task of end-to-end referential bridging resolution in English. Our bridging resolu

LoMo: Local Modality Substitution for Deeper Vision-Language Fusion

SafetyDGX agent

arXiv:2605.30265v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) have achieved substantial progress across a wide range of understanding and reasoning tasks, driven by large-scale image

Long-Context Modeling with Dynamic Hierarchical Sparse Attention for Memory-Constrained LLM Inference

Model ReleasesDGX agent

arXiv:2510.24606v2 Announce Type: replace Abstract: The quadratic cost of attention limits the scalability of long-context LLMs, especially under limited hardware memory budgets. While attention is of

← Previous
1…5354555657…129
Next →