AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlog
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
Model Releases

CorrSteer: Generation-Time LLM Steering via Correlated Sparse Autoencoder Features

DGX agent

arXiv:2508.12535v3 Announce Type: replace Abstract: Sparse Autoencoders (SAEs) can extract interpretable features from large language models (LLMs) without supervision. However, their effectiveness in

model-releasesarxiv-cs-cl
5 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

CoSpaDi: Compressing LLMs via Calibration-Guided Sparse Dictionary Learning

DGX agent

arXiv:2509.22075v5 Announce Type: replace Abstract: Post-training compression of large language models (LLMs) often relies on low-rank weight approximations that represent each column of the weight ma

model-releasesarxiv-cs-cl
5 May 2026
Research

Counting as a minimal probe of language model reliability

DGX agent

arXiv:2605.02028v1 Announce Type: new Abstract: Large language models perform strongly on benchmarks in mathematical reasoning, coding and document analysis, suggesting a broad ability to follow instr

researcharxiv-cs-cl
5 May 2026
Model Releases

CP-SynC: Multi-Agent Zero-Shot Constraint Modeling in MiniZinc with Synthesized Checkers

DGX agent

arXiv:2605.01675v1 Announce Type: cross Abstract: Constraint Programming (CP) is a powerful paradigm for solving combinatorial problems, yet translating natural language problem descriptions into exec

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Creating and Evaluating Figurative Language Dataset for Sindhi

DGX agent

arXiv:2605.01323v1 Announce Type: new Abstract: In this article, we introduce SiNFluD, a novel benchmark dataset for Sindhi figurative language classification. We first collect raw text from various b

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

CyclicJudge: Mitigating Judge Bias Efficiently in LLM-based Evaluation

DGX agent

arXiv:2603.01865v3 Announce Type: replace Abstract: LLM-as-judge evaluation has become standard practice for open-ended model assessment; however, judges exhibit systematic biases that cannot be avera

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Decoding-Time Debiasing via Process Reward Models: From Controlled Fill-in to Open-Ended Generation

DGX agent

arXiv:2605.02348v1 Announce Type: new Abstract: Large language models pick up social biases from the data they are trained on and carry those biases into downstream applications, often reinforcing ste

model-releasesarxiv-cs-cl
5 May 2026
Hardware

DELTA: Dynamic Layer-Aware Token Attention for Efficient Long-Context Reasoning

DGX agent

arXiv:2510.09883v2 Announce Type: replace Abstract: Large reasoning models (LRMs) achieve state-of-the-art performance on challenging benchmarks by generating long chains of intermediate steps, but th

hardwarearxiv-cs-cl
5 May 2026
Model Releases

Democratizing the medieval English legal tradition

DGX agent

arXiv:2605.00977v1 Announce Type: cross Abstract: The record of the beginning of the most widespread legal system in the world is contained in millions of pages of handwritten text. Most of the record

model-releasesarxiv-cs-cl
5 May 2026
Research

Dependency Parsing Across the Resource Spectrum: Evaluating Architectures on High and Low-Resource Languages

DGX agent

arXiv:2605.02608v1 Announce Type: new Abstract: Transformer-based models achieve state-of-the-art dependency parsing for high-resource languages, yet their advantage over simpler architectures in low-

researcharxiv-cs-cl
5 May 2026
Research

DIAGRAMS: A Review Framework for Reasoning-Level Attribution in Diagram QA

DGX agent

arXiv:2605.00905v1 Announce Type: new Abstract: Diagram question answering (Diagram QA) requires reasoning-level attribution that links each question-answer pair to all visual regions needed to derive

researcharxiv-cs-cl
5 May 2026
Safety

Do Large Language Models Plan Answer Positions? Position Bias in Multiple-Choice Question Generation

DGX agent

arXiv:2605.01846v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used to generate multiple-choice questions (MCQs), where correct answers should ideally be uniformly distr

safetyarxiv-cs-cl
5 May 2026
Model Releases

EditPropBench: Measuring Factual Edit Propagation in Scientific Manuscripts

DGX agent

arXiv:2605.02083v1 Announce Type: new Abstract: Local factual edits in scientific manuscripts often create non-local revision obligations. If a dataset changes from 215 to 80 documents, claims such as

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Efficient Reasoning with Hidden Thinking

DGX agent

arXiv:2501.19201v2 Announce Type: replace Abstract: Chain-of-Thought (CoT) reasoning has become a powerful framework for improving complex problem-solving capabilities in Multimodal Large Language Mod

model-releasesarxiv-cs-cl
5 May 2026
Tutorials

EGAD: Entropy-Guided Adaptive Distillation for Token-Level Knowledge Transfer

DGX agent

arXiv:2605.01732v1 Announce Type: new Abstract: Large language models (LLMs) have achieved remarkable performance across diverse domains, yet their enormous computational and memory requirements hinde

tutorialsarxiv-cs-cl
5 May 2026
Model Releases

Embedding-based In-Context Prompt Training for Enhancing LLMs as Text Encoders

DGX agent

arXiv:2605.01372v1 Announce Type: new Abstract: Large language models (LLMs) have been widely explored for embedding generation. While recent studies show that in-context learning (ICL) effectively en

model-releasesarxiv-cs-cl
5 May 2026
Local Ai

Energy-Based Constraint Networks: Learning Structural Coherence Across Modalities

DGX agent

arXiv:2605.00960v1 Announce Type: cross Abstract: We introduce energy-based constraint networks -- a modality-agnostic architecture that learns structural coherence from contrastive pairs. The system

local-aiarxiv-cs-cl
5 May 2026
Model Releases

Enhanced LLM Reasoning by Optimizing Reward Functions with Search-Driven Reinforcement Learning

DGX agent

arXiv:2605.02073v1 Announce Type: new Abstract: Mathematical reasoning is a key benchmark for large language models. Reinforcement learning is a standard post-training mechanism for improving the reas

model-releasesarxiv-cs-cl
5 May 2026
Research

Enhancing Game Review Sentiment Classification on Steam Platform with Attention-Based BiLSTM

DGX agent

arXiv:2605.01315v1 Announce Type: new Abstract: This paper investigates sentiment classification of Steam game reviews using an attention-based Bidirectional Long Short-Term Memory (BiLSTM) model. Usi

researcharxiv-cs-cl
5 May 2026
Model Releases

Enhancing Judgment Document Generation via Agentic Legal Information Collection and Rubric-Guided Optimization

DGX agent

arXiv:2605.02011v1 Announce Type: new Abstract: Automating the drafting of judgment documents is pivotal to judicial efficiency, yet it remains challenging due to the dual requirements of comprehensiv

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Extracting memorized pieces of (copyrighted) books from open-weight language models

DGX agent

arXiv:2505.12546v5 Announce Type: replace Abstract: Plaintiffs and defendants in copyright lawsuits over generative AI often make sweeping, opposing claims about the extent to which large language mod

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Feedback-Normalized Developer Memory for Reinforcement-Learning Coding Agents: A Safety-Gated MCP Architecture

DGX agent

arXiv:2605.01567v1 Announce Type: cross Abstract: Large language model (LLM) coding agents increasingly operate over repositories, terminals, tests, and execution traces across long software-engineeri

model-releasesarxiv-cs-cl
5 May 2026
Research

Fight Poison with Poison: Enhancing Robustness in Few-shot Machine-Generated Text Detection with Adversarial Training

DGX agent

arXiv:2605.02374v1 Announce Type: cross Abstract: Machine-generated text (MGT) detection is critical for regulating online information ecosystems, yet existing detectors often underperform in few-shot

researcharxiv-cs-cl
5 May 2026
Model Releases

Fin-PRM: A Domain-Specialized Process Reward Model for Financial Reasoning in Large Language Models

DGX agent

arXiv:2508.15202v2 Announce Type: replace Abstract: Process Reward Models (PRMs) supervise intermediate reasoning steps in large language models (LLMs), but existing PRMs are mainly trained on general

model-releasesarxiv-cs-cl
5 May 2026
Research

Fine-Tuning Pre-Trained Code Models for AI-Generated Code Detection

DGX agent

arXiv:2605.01596v1 Announce Type: new Abstract: This paper describes the system submitted by team extbf{Archaeology} to SemEval-2026 Task~13 on AI-generated code detection. The shared task consists of

researcharxiv-cs-cl
5 May 2026
Model Releases

Flexi-LoRA with Input-Adaptive Ranks: Efficient Finetuning for Speech and Reasoning Tasks

DGX agent

arXiv:2605.01959v1 Announce Type: cross Abstract: Parameter-efficient fine-tuning methods like Low-Rank Adaptation (LoRA) have become essential for deploying large language models, yet their static pa

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

FlexSQL: Flexible Exploration and Execution Make Better Text-to-SQL Agents

DGX agent

arXiv:2605.02815v1 Announce Type: new Abstract: Text-to-SQL over large analytical databases requires navigating complex schemas, resolving ambiguous queries, and grounding decisions in actual data. Mo

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Focus on the Core: Empowering Diffusion Large Language Models by Self-Contrast

DGX agent

arXiv:2605.01373v1 Announce Type: new Abstract: The iterative denoising paradigm of Diffusion Large Language Models (DLMs) endows them with a distinct advantage in global context modeling. However, cu

model-releasesarxiv-cs-cl
5 May 2026
Safety

Foundation Models to Unlock Real-World Evidence from Nationwide Medical Claims

DGX agent

arXiv:2605.02740v1 Announce Type: cross Abstract: Evidence derived from large-scale real-world data (RWD) is increasingly informing regulatory evaluation and healthcare decision-making. Administrative

safetyarxiv-cs-cl
5 May 2026
Model Releases

FT-RAG: A Fine-grained Retrieval-Augmented Generation Framework for Complex Table Reasoning

DGX agent

arXiv:2605.01495v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) enhances Large Language Models (LLMs) by grounding responses in external knowledge during inference. However, conve

model-releasesarxiv-cs-cl
5 May 2026
Research

FunFuzz: An LLM-Powered Evolutionary Fuzzing Framework

DGX agent

arXiv:2605.02789v1 Announce Type: cross Abstract: Modern fuzzers increasingly use Large Language Models (LLMs) to generate structured inputs, but LLM-driven fuzzing is sensitive to prompt initializati

researcharxiv-cs-cl
5 May 2026
Research

Fuzzy Fingerprinting Encoder Pre-trained Language Models for Emotion Recognition in Conversations: Human Assessment and Validity Study

DGX agent

arXiv:2605.02665v1 Announce Type: new Abstract: In Emotion Recognition in Conversations (ERC), model decisions should align with nuanced human perception and ideally provide insights on the classifica

researcharxiv-cs-cl
5 May 2026
Research

Generative Interfaces for Language Models

DGX agent

arXiv:2508.19227v3 Announce Type: replace Abstract: Large language models (LLMs) are increasingly seen as assistants, copilots, and consultants, capable of supporting a wide range of tasks through nat

researcharxiv-cs-cl
5 May 2026
Tutorials

GIFT: Guided Fine-Tuning and Transfer for Enhancing Instruction-Tuned Language Models

DGX agent

arXiv:2605.01256v1 Announce Type: new Abstract: A promising paradigm for adapting instruction-tuned language models is to learn task-specific updates on a pretrained base model and subsequently merge

tutorialsarxiv-cs-cl
5 May 2026
Model Releases

GR-Ben: A General Reasoning Benchmark for Evaluating Process Reward Models

DGX agent

arXiv:2605.01203v1 Announce Type: cross Abstract: Currently, process reward models (PRMs) have exhibited remarkable potential for test-time scaling. Since large language models (LLMs) regularly genera

model-releasesarxiv-cs-cl
5 May 2026
Agents

GRAIL: A Deep-Granularity Hybrid Resonance Framework for Real-Time Agent Discovery via SLM-Enhanced Indexing

DGX agent

arXiv:2605.02489v1 Announce Type: cross Abstract: As the ecosystem of Large Language Model (LLM)-based agents expands rapidly, efficient and accurate Agent Discovery becomes a critical bottleneck for

agentsarxiv-cs-cl
5 May 2026
Applications

Graph Query Generation with Constraint-guided Large Language Agents

DGX agent

arXiv:2605.00845v1 Announce Type: cross Abstract: Knowledge Graph Question Answering (KGQA) has advanced through structured query generation, yet most efforts target RDF/SPARQL, leaving Cypher and pro

applicationsarxiv-cs-cl
5 May 2026
Research

GRAVITY: Architecture-Agnostic Structured Anchoring for Long-Horizon Conversational Memory

DGX agent

arXiv:2605.01688v1 Announce Type: new Abstract: Long-horizon conversational agents rely on memory systems with increasingly sophisticated retrieval mechanisms. However, retrieved fragments are typical

researcharxiv-cs-cl
5 May 2026
Model Releases

Growing Transformers: Modular Composition and Layer-wise Expansion on a Frozen Substrate

DGX agent

arXiv:2507.07129v3 Announce Type: replace-cross Abstract: We study a constrained training regime for decoder-only Transformers in which the token interface is fixed, previously trained dense blocks ar

model-releasesarxiv-cs-cl
5 May 2026
Applications

H-Probes: Extracting Hierarchical Structures From Latent Representations of Language Models

DGX agent

arXiv:2605.00847v1 Announce Type: new Abstract: Representing and navigating hierarchy is a fundamental primitive of reasoning. Large language models have demonstrated proficiency in a wide variety of

applicationsarxiv-cs-cl
5 May 2026
Research

Hallucination Detection in LLMs with Topological Divergence on Attention Graphs

DGX agent

arXiv:2504.10063v4 Announce Type: replace Abstract: Hallucination, i.e., generating factually incorrect content, remains a critical challenge for large language models (LLMs). We introduce TOHA, a TOp

researcharxiv-cs-cl
5 May 2026
Agents

Hallucinations Undermine Trust; Metacognition is a Way Forward

DGX agent

arXiv:2605.01428v1 Announce Type: new Abstract: Despite significant strides in factual reliability, errors -- often termed hallucinations -- remain a major concern for generative AI, especially as LLM

agentsarxiv-cs-cl
5 May 2026
Model Releases

HalluScan: A Systematic Benchmark for Detecting and Mitigating Hallucinations in Instruction-Following LLMs

DGX agent

arXiv:2605.02443v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated remarkable capabilities across diverse natural language processing tasks, yet they remain susceptible to

model-releasesarxiv-cs-cl
5 May 2026
Safety

HeteroRAG: A Heterogeneous Retrieval-Augmented Generation Framework for Medical Vision Language Tasks

DGX agent

arXiv:2508.12778v2 Announce Type: replace Abstract: Medical large vision-language Models (Med-LVLMs) have shown promise in clinical applications but suffer from factual inaccuracies and unreliable out

safetyarxiv-cs-cl
5 May 2026
Research

Hey, That's My Data! Token-Only Dataset Inference in Large Language Models

DGX agent

arXiv:2506.06057v2 Announce Type: replace Abstract: Large Language Models (LLMs) rely on massive training datasets, often including proprietary data, which raises concerns about unauthorized usage and

researcharxiv-cs-cl
5 May 2026
Research

How Good is Your Wikipedia? Auditing Data Quality for Low-resource and Multilingual NLP

DGX agent

arXiv:2411.05527v3 Announce Type: replace Abstract: Wikipedia's perceived high quality and broad language coverage have established it as a fundamental resource in NLP. However, in recent years, such

researcharxiv-cs-cl
5 May 2026
Research

How Prompts Move Language Model Behavior: Frames, Salience, and Construal as Semantic Control

DGX agent

arXiv:2512.12688v3 Announce Type: replace-cross Abstract: Prompt engineering is widely used to shape large language model behavior, yet it is often treated as a practical heuristic rather than as a fo

researcharxiv-cs-cl
5 May 2026
Model Releases

How Well Can We Decode Vowels from Auditory EEG -- A Rigorous Cross-Subject Benchmark with Honest Assessment

DGX agent

arXiv:2605.00865v1 Announce Type: cross Abstract: EEG based phoneme decoding is promising for brain computer interfaces, but many prior studies rely on within subject evaluation, small cohorts, or wea

model-releasesarxiv-cs-cl
5 May 2026
← Previous
1…109110111112113…161
Next →