AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,646 results
Local Ai

Dissecting Failure Dynamics in Large Language Model Reasoning

DGX agent

arXiv:2604.14528v1 Announce Type: cross Abstract: Large Language Models (LLMs) achieve strong performance through extended inference-time deliberation, yet how their reasoning failures arise remains p

local-aiarxiv-cs-cl
17 Apr 2026
Model Releases
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Dive into Claude Code: The Design Space of Today's and Future AI Agent Systems

DGX agent

arXiv:2604.14228v1 Announce Type: cross Abstract: Claude Code is an agentic coding tool that can run shell commands, edit files, and call external services on behalf of the user. This study describes

model-releasesarxiv-cs-cl
17 Apr 2026
Applications

Domain Fine-Tuning FinBERT on Finnish Histopathological Reports: Train-Time Signals and Downstream Correlations

DGX agent

arXiv:2604.14815v1 Announce Type: new Abstract: In NLP classification tasks where little labeled data exists, domain fine-tuning of transformer models on unlabeled data is an established approach. In

applicationsarxiv-cs-cl
17 Apr 2026
Model Releases

Don't Retrieve, Navigate: Distilling Enterprise Knowledge into Navigable Agent Skills for QA and RAG

DGX agent

arXiv:2604.14572v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) grounds LLM responses in external evidence but treats the model as a passive consumer of search results: it never

model-releasesarxiv-cs-cl
17 Apr 2026
Research

DySCO: Dynamic Attention-Scaling Decoding for Long-Context Language Models

DGX agent

arXiv:2602.22175v2 Announce Type: replace Abstract: Understanding and reasoning over long contexts is a crucial capability for language models (LMs). Although recent models support increasingly long c

researcharxiv-cs-cl
17 Apr 2026
Research

Enhancing Linguistic Competence of Language Models through Pre-training with Language Learning Tasks

DGX agent

arXiv:2601.03448v2 Announce Type: replace Abstract: Language models (LMs) are pre-trained on raw text datasets to generate text sequences token-by-token. While this approach facilitates the learning o

researcharxiv-cs-cl
17 Apr 2026
Model Releases

EuropeMedQA Study Protocol: A Multilingual, Multimodal Medical Examination Dataset for Language Model Evaluation

DGX agent

arXiv:2604.14306v1 Announce Type: new Abstract: While Large Language Models (LLMs) have demonstrated high proficiency on English-centric medical examinations, their performance often declines when fac

model-releasesarxiv-cs-cl
17 Apr 2026
Model Releases

EviSearch: A Human in the Loop System for Extracting and Auditing Clinical Evidence for Systematic Reviews

DGX agent

arXiv:2604.14165v1 Announce Type: new Abstract: We present EviSearch, a multi-agent extraction system that automates the creation of ontology-aligned clinical evidence tables directly from native tria

model-releasesarxiv-cs-cl
17 Apr 2026
Research

Evolving Beyond Snapshots: Harmonizing Structure and Sequence via Entity State Tuning for Temporal Knowledge Graph Forecasting

DGX agent

arXiv:2602.12389v3 Announce Type: replace-cross Abstract: Temporal knowledge graph (TKG) forecasting requires predicting future facts by jointly modeling structural dependencies within each snapshot a

researcharxiv-cs-cl
17 Apr 2026
Research

Explain the Flag: Contextualizing Hate Speech Beyond Censorship

DGX agent

arXiv:2604.14970v1 Announce Type: new Abstract: Hate, derogatory, and offensive speech remains a persistent challenge in online platforms and public discourse. While automated detection systems are wi

researcharxiv-cs-cl
17 Apr 2026
Model Releases

Exploring and Testing Skill-Based Behavioral Profile Annotation: Human Operability and LLM Feasibility under Schema-Guided Execution

DGX agent

arXiv:2604.14843v1 Announce Type: new Abstract: Behavioral Profile (BP) annotation is difficult to automate because it requires simultaneous coding across multiple linguistic dimensions. We treat BP a

model-releasesarxiv-cs-cl
17 Apr 2026
Research

Fabricator or dynamic translator?

DGX agent

arXiv:2604.15165v1 Announce Type: new Abstract: LLMs are proving to be adept at machine translation although due to their generative nature they may at times overgenerate in various ways. These overge

researcharxiv-cs-cl
17 Apr 2026
Model Releases

Fact4ac at the Financial Misinformation Detection Challenge Task: Reference-Free Financial Misinformation Detection via Fine-Tuning and Few-Shot Prompting of Large Language Models

DGX agent

arXiv:2604.14640v1 Announce Type: new Abstract: The proliferation of financial misinformation poses a severe threat to market stability and investor trust, misleading market behavior and creating crit

model-releasesarxiv-cs-cl
17 Apr 2026
Research

Faithfulness Serum: Mitigating the Faithfulness Gap in Textual Explanations of LLM Decisions via Attribution Guidance

DGX agent

arXiv:2604.14325v1 Announce Type: new Abstract: Large language models (LLMs) achieve strong performance and have revolutionized NLP, but their lack of explainability keeps them treated as black boxes,

researcharxiv-cs-cl
17 Apr 2026
Research

Feedback Adaptation for Retrieval-Augmented Generation

DGX agent

arXiv:2604.06647v2 Announce Type: replace Abstract: Retrieval-Augmented Generation (RAG) systems are typically evaluated under static assumptions, despite being frequently corrected through user or ex

researcharxiv-cs-cl
17 Apr 2026
Safety

Filling in the Mechanisms: How do LMs Learn Filler-Gap Dependencies under Developmental Constraints?

DGX agent

arXiv:2604.14459v1 Announce Type: new Abstract: For humans, filler-gap dependencies require a shared representation across different syntactic constructions. Although causal analyses suggest this may

safetyarxiv-cs-cl
17 Apr 2026
Applications

From Black Box to Glass Box: Cross-Model ASR Disagreement to Prioto Review in Ambient AI Scribe Documentation

DGX agent

arXiv:2604.14152v1 Announce Type: cross Abstract: Ambient AI 'scribe' systems promise to reduce clinical documentation burden, but automatic speech recognition (ASR) errors can remain unnoticed withou

applicationsarxiv-cs-cl
17 Apr 2026
Safety

From Plausible to Causal: Counterfactual Semantics for Policy Evaluation in Simulated Online Communities

DGX agent

arXiv:2604.03920v2 Announce Type: replace Abstract: LLM-based social simulations can generate believable community interactions, enabling ``policy wind tunnels'' where governance interventions are tes

safetyarxiv-cs-cl
17 Apr 2026
Tutorials

From Procedural Skills to Strategy Genes: Towards Experience-Driven Test-Time Evolution

DGX agent

arXiv:2604.15097v1 Announce Type: cross Abstract: This beta technical report asks how reusable experience should be represented so that it can function as effective test-time control and as a substrat

tutorialsarxiv-cs-cl
17 Apr 2026
Research

From Reactive to Proactive: Assessing the Proactivity of Voice Agents via ProVoice-Bench

DGX agent

arXiv:2604.15037v1 Announce Type: cross Abstract: Recent advancements in LLM agents are gradually shifting from reactive, text-based paradigms toward proactive, multimodal interaction. However, existi

researcharxiv-cs-cl
17 Apr 2026
Research

From Tokens to Steps: Verification-Aware Speculative Decoding for Efficient Multi-Step Reasoning

DGX agent

arXiv:2604.15244v1 Announce Type: new Abstract: Speculative decoding (SD) accelerates large language model inference by allowing a lightweight draft model to propose outputs that a stronger target mod

researcharxiv-cs-cl
17 Apr 2026
Research

Generating Concept Lexicalizations via Dictionary-Based Cross-Lingual Sense Projection

DGX agent

arXiv:2604.14397v1 Announce Type: new Abstract: We study the task of automatically expanding WordNet-style lexical resources to new languages through sense generation. We generate senses by associatin

researcharxiv-cs-cl
17 Apr 2026
Research

Grading the Unspoken: Evaluating Tacit Reasoning in Quantum Field Theory and String Theory with LLMs

DGX agent

arXiv:2604.14188v1 Announce Type: cross Abstract: Large language models have demonstrated impressive performance across many domains of mathematics and physics. One natural question is whether such mo

researcharxiv-cs-cl
17 Apr 2026
Research

Graph-Based Alternatives to LLMs for Human Simulation

DGX agent

arXiv:2511.02135v2 Announce Type: replace Abstract: Large language models (LLMs) have become a popular approach for simulating human behaviors, yet it remains unclear if LLMs are necessary for all sim

researcharxiv-cs-cl
17 Apr 2026
Applications

HARNESS: Lightweight Distilled Arabic Speech Foundation Models

DGX agent

arXiv:2604.14186v1 Announce Type: cross Abstract: Large self-supervised speech (SSL) models achieve strong downstream performance, but their size limits deployment in resource-constrained settings. We

applicationsarxiv-cs-cl
17 Apr 2026
Hardware

HERMES: KV Cache as Hierarchical Memory for Efficient Streaming Video Understanding

DGX agent

arXiv:2601.14724v3 Announce Type: replace-cross Abstract: Recent advancements in Multimodal Large Language Models (MLLMs) have demonstrated significant improvement in offline video understanding. Howe

hardwarearxiv-cs-cl
17 Apr 2026
Safety

Hierarchical Retrieval Augmented Generation for Adversarial Technique Annotation in Cyber Threat Intelligence Text

DGX agent

arXiv:2604.14166v1 Announce Type: new Abstract: Mapping Cyber Threat Intelligence (CTI) text to MITRE ATT&CK technique IDs is a critical task for understanding adversary behaviors and automating threa

safetyarxiv-cs-cl
17 Apr 2026
Research

Hierarchical Semantic Retrieval with Cobweb

DGX agent

arXiv:2510.02539v2 Announce Type: replace Abstract: Neural document retrieval often treats a corpus as a flat cloud of vectors scored at a single granularity, leaving corpus structure underused and ex

researcharxiv-cs-cl
17 Apr 2026
Model Releases

Hierarchical vs. Flat Iteration in Shared-Weight Transformers

DGX agent

arXiv:2604.14442v1 Announce Type: new Abstract: We present an empirical study of whether hierarchically structured, shared-weight recurrence can match the representational quality of independent-layer

model-releasesarxiv-cs-cl
17 Apr 2026
Research

How Retrieved Context Shapes Internal Representations in RAG

DGX agent

arXiv:2602.20091v2 Announce Type: replace Abstract: Retrieval-augmented generation (RAG) enhances large language models (LLMs) by conditioning generation on retrieved external documents, but the effec

researcharxiv-cs-cl
17 Apr 2026
Tutorials

How to Fine-Tune a Reasoning Model? A Teacher-Student Cooperation Framework to Synthesize Student-Consistent SFT Data

DGX agent

arXiv:2604.14164v1 Announce Type: new Abstract: A widely adopted strategy for model enhancement is to use synthetic data generated by a stronger model for supervised fine-tuning (SFT). However, for em

tutorialsarxiv-cs-cl
17 Apr 2026
Local Ai

HUOZIIME: An On-Device LLM-enhanced Input Method for Deep Personalization

DGX agent

arXiv:2604.14159v1 Announce Type: new Abstract: Mobile input method editors (IMEs) are the primary interface for text input, yet they remain constrained to manual typing and struggle to produce person

local-aiarxiv-cs-cl
17 Apr 2026
Tutorials

Hybrid Decision Making via Conformal VLM-generated Guidance

DGX agent

arXiv:2604.14980v1 Announce Type: cross Abstract: Building on recent advances in AI, hybrid decision making (HDM) holds the promise of improving human decision quality and reducing cognitive load. We

tutorialsarxiv-cs-cl
17 Apr 2026
Agents

IE as Cache: Information Extraction Enhanced Agentic Reasoning

DGX agent

arXiv:2604.14930v1 Announce Type: new Abstract: Information Extraction aims to distill structured, decision-relevant information from unstructured text, serving as a foundation for downstream understa

agentsarxiv-cs-cl
17 Apr 2026
Model Releases

IF-CRITIC: Towards a Fine-Grained LLM Critic for Instruction-Following Evaluation

DGX agent

arXiv:2511.01014v3 Announce Type: replace Abstract: Instruction-following is a fundamental ability of Large Language Models (LLMs), requiring their generated outputs to follow multiple constraints imp

model-releasesarxiv-cs-cl
17 Apr 2026
Model Releases

IF-RewardBench: Benchmarking Judge Models for Instruction-Following Evaluation

DGX agent

arXiv:2603.04738v2 Announce Type: replace Abstract: Instruction-following is a foundational capability of large language models (LLMs), with its improvement hinging on scalable and accurate feedback f

model-releasesarxiv-cs-cl
17 Apr 2026
Safety

IG-Search: Step-Level Information Gain Rewards for Search-Augmented Reasoning

DGX agent

arXiv:2604.15148v1 Announce Type: cross Abstract: Reinforcement learning has emerged as an effective paradigm for training large language models to perform search-augmented reasoning. However, existin

safetyarxiv-cs-cl
17 Apr 2026
Model Releases

Improving Language Models with Intentional Analysis

DGX agent

arXiv:2502.04689v4 Announce Type: replace Abstract: Intent, a critical cognitive notion and mental state, is ubiquitous in human communication and problem-solving. Accurately understanding the underly

model-releasesarxiv-cs-cl
17 Apr 2026
Model Releases

In Context Learning and Reasoning for Symbolic Regression with Large Language Models

DGX agent

arXiv:2410.17448v3 Announce Type: replace Abstract: Large Language Models (LLMs) are transformer-based machine learning models that have shown remarkable performance in tasks for which they were not e

model-releasesarxiv-cs-cl
17 Apr 2026
Model Releases

Internal Knowledge Without External Expression: Probing the Generalization Boundary of a Classical Chinese Language Model

DGX agent

arXiv:2604.14180v1 Announce Type: new Abstract: We train a 318M-parameter Transformer language model from scratch on a curated corpus of 1.56 billion tokens of pure Classical Chinese, with zero Englis

model-releasesarxiv-cs-cl
17 Apr 2026
Safety

IROSA: Interactive Robot Skill Adaptation using Natural Language

DGX agent

arXiv:2603.03897v3 Announce Type: replace-cross Abstract: Foundation models have demonstrated impressive capabilities across diverse domains, while imitation learning provides principled methods for r

safetyarxiv-cs-cl
17 Apr 2026
Applications

IUQ: Interrogative Uncertainty Quantification for Long-Form Large Language Model Generation

DGX agent

arXiv:2604.15109v1 Announce Type: new Abstract: Despite the rapid advancement of Large Language Models (LLMs), uncertainty quantification in LLM generation is a persistent challenge. Although recent a

applicationsarxiv-cs-cl
17 Apr 2026
Research

Just Pass Twice: Efficient Token Classification with LLMs for Zero-Shot NER

DGX agent

arXiv:2604.05158v2 Announce Type: replace Abstract: Large language models encode extensive world knowledge valuable for zero-shot named entity recognition. However, their causal attention mechanism, w

researcharxiv-cs-cl
17 Apr 2026
Model Releases

Knowing When Not to Answer: Evaluating Abstention in Multimodal Reasoning Systems

DGX agent

arXiv:2604.14799v1 Announce Type: new Abstract: Effective abstention (EA), recognizing evidence insufficiency and refraining from answering, is critical for reliable multimodal systems. Yet existing e

model-releasesarxiv-cs-cl
17 Apr 2026
Tutorials

KnowRL: Exploring Knowledgeable Reinforcement Learning for Factuality

DGX agent

arXiv:2506.19807v4 Announce Type: replace-cross Abstract: Large Language Models (LLMs), particularly slow-thinking models, often exhibit severe hallucination, outputting incorrect content due to an in

tutorialsarxiv-cs-cl
17 Apr 2026
Safety

Language Model as Planner and Formalizer under Constraints

DGX agent

arXiv:2510.05486v2 Announce Type: replace Abstract: LLMs have been widely used in planning, either as planners to generate action sequences end-to-end, or as formalizers to represent the planning doma

safetyarxiv-cs-cl
17 Apr 2026
Applications

Language Model Fine-Tuning on Scaled Survey Data for Predicting Distributions of Public Opinions

DGX agent

arXiv:2502.16761v2 Announce Type: replace Abstract: Large language models (LLMs) present novel opportunities in public opinion research by predicting survey responses in advance during the early stage

applicationsarxiv-cs-cl
17 Apr 2026
Safety

Language of Thought Shapes Output Diversity in Large Language Models

DGX agent

arXiv:2601.11227v2 Announce Type: replace Abstract: Output diversity is crucial for Large Language Models as it underpins pluralism and creativity. In this work, we reveal that controlling the languag

safetyarxiv-cs-cl
17 Apr 2026
← Previous
1…142143144145146…160
Next →