AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
Model Releases

PilotRL: Training Language Model Agents via Global Planning-Guided Progressive Reinforcement Learning

DGX agent

arXiv:2508.00344v5 Announce Type: replace Abstract: Large Language Models (LLMs) have shown remarkable advancements in tackling agent-oriented tasks. Despite their potential, existing work faces chall

model-releasesarxiv-cs-cl
29 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Polistemics: Evaluating LLMs as Information Mediators in Politics & Elections

DGX agent

arXiv:2607.25953v1 Announce Type: new Abstract: As LLMs increasingly mediate the political information citizens rely on, there is still no standardized way to assess whether they do so responsibly. We

model-releasesarxiv-cs-cl
29 Jul 2026
Safety

Ranked by Position: Order Sensitivity as an Exploitable Attack Surface in LLM Listwise Recommenders

DGX agent

arXiv:2607.24869v1 Announce Type: cross Abstract: Large language models (LLMs) used as listwise rerankers in recommendation systems suffer from position bias when serializing candidate sets into promp

safetyarxiv-cs-cl
29 Jul 2026
Model Releases

Research Report on Noise-Shaped One-Bit Coefficients in Discrete Polynomial Fourier Extension

DGX agent

arXiv:2607.24868v1 Announce Type: new Abstract: This report studies noise-shaped one-bit coefficients in normalized discrete polynomial Fourier extension. For first-order Sigma-Delta quantization, the

model-releasesarxiv-cs-cl
29 Jul 2026
Research

Retrieval, not hallucinations, will be the limiting factor for LLM-based clinical AI tools

DGX agent

arXiv:2607.24793v1 Announce Type: cross Abstract: Discussions around large language model (LLM) errors in clinical artificial intelligence (AI) generally center around precision errors like hallucinat

researcharxiv-cs-cl
29 Jul 2026
Model Releases

RSIBench-Data: Benchmarking Data-Centric Research for Recursive Self-Improvement

DGX agent

arXiv:2607.25886v1 Announce Type: cross Abstract: Recursive self-improvement requires turning evidence of model failures into better models. Data-centric post-training research entails diagnosing capa

model-releasesarxiv-cs-cl
29 Jul 2026
Model Releases

SAGE: Stochastic Prompt Optimization via Agent-Guided Exploration

DGX agent

arXiv:2606.18902v2 Announce Type: replace Abstract: Context engineering has emerged as a primary lever for improving AI systems without parameter updates. Recent work showing that textual gradients do

model-releasesarxiv-cs-cl
29 Jul 2026
Research

SciClaimSeekers at CheckThat! 2026: Retrieving Scientific Sources for Social Media Claims with LLM Reranking

DGX agent

arXiv:2607.24803v1 Announce Type: cross Abstract: Scientific claims often spread on social media faster than they can be verified, while posts rarely link to the original scholarly sources. To tackle

researcharxiv-cs-cl
29 Jul 2026
Model Releases

Shieldstral

DGX agent

arXiv:2607.25857v1 Announce Type: new Abstract: We introduce Shieldstral, a 3B-parameter policy-adaptive multimodal safety classifier that matches or outperforms models nearly 7imes its size on text s

model-releasesarxiv-cs-cl
29 Jul 2026
Agents

SourceMinds at CheckThat! 2026: NLI-Grounded Citation Auditing in a Multi-Agent Pipeline for Full Fact-Checking Article Generation

DGX agent

arXiv:2607.24802v1 Announce Type: cross Abstract: This paper presents our system for Task 3 of the CLEF 2026 CheckThat! Lab, which focuses on generating full fact-checking articles from claims, veraci

agentsarxiv-cs-cl
29 Jul 2026
Model Releases

SpeechLLM Meets Federated Learning for End-to-End ASR: English and Italian Case Studies

DGX agent

arXiv:2607.25716v1 Announce Type: new Abstract: Federated learning (FL) enables privacy-preserving training of automatic speech recognition (ASR) systems across distributed data sources, yet its appli

model-releasesarxiv-cs-cl
29 Jul 2026
Tutorials

Temporal-Distance JEPA: Plan-Aware Representation Learning for Latent World Model Predictive Control

DGX agent

arXiv:2607.25337v1 Announce Type: new Abstract: Joint-Embedding Predictive Architectures (JEPAs) learn world models by predicting in representation space rather than reconstructing pixels, making them

tutorialsarxiv-cs-cl
29 Jul 2026
Model Releases

TimeCapsule: Generative Hallucination as a Method for Historical Sensemaking

DGX agent

arXiv:2607.24750v1 Announce Type: new Abstract: Large Language Models (LLMs) are temporally overexposed: trained on vast contemporary corpora, they encode present-day concepts that make them unreliabl

model-releasesarxiv-cs-cl
29 Jul 2026
Research

Toward a systematic method for identifying language areas

DGX agent

arXiv:2607.25305v1 Announce Type: new Abstract: Macroareas are geographical areas used in typological research for grouping variables of interest. In linguistic typology, languages in a given macroare

researcharxiv-cs-cl
29 Jul 2026
Model Releases

UniMem: Complementary Episodic-to-Parametric Memory for Boundary-Agnostic Task Streams

DGX agent

arXiv:2607.26017v1 Announce Type: new Abstract: Memory is essential for LLM agents to accumulate task experience and reuse task-specific execution strategies. However, real-world deployment over bound

model-releasesarxiv-cs-cl
29 Jul 2026
Research

VisRAG2.0: Mitigating Visual Hallucinations via Evidence-Guided Multi-Image Reasoning in Visual Retrieval-Augmented Generation

DGX agent

arXiv:2510.09733v2 Announce Type: replace Abstract: Visual Retrieval-Augmented Generation (VRAG) has emerged as a promising paradigm for equipping Vision-Language Models (VLMs) with external visual ev

researcharxiv-cs-cl
29 Jul 2026
Tutorials

VisualPatchWorld: Code World Models as Latent Structured Representations for Planning

DGX agent

arXiv:2607.25236v1 Announce Type: new Abstract: Different research lines use the term world model in different ways, yet they share a common aim: to capture how the world evolves under action in a for

tutorialsarxiv-cs-cl
29 Jul 2026
Applications

When Algorithms Meet Artists: Semantic Compression and Stake-holder Marginalisation in Public AI-Art Discourse (2013-2025)

DGX agent

arXiv:2508.03037v5 Announce Type: replace Abstract: Artists occupy a paradoxical position in generative AI. Their own work trains models that now compete with them, replicate their styles, and reshape

applicationsarxiv-cs-cl
29 Jul 2026
Model Releases

WorkSurface-Bench: Benchmarking Enterprise Agents on Multi-Surface Knowledge Routing

DGX agent

arXiv:2607.25765v1 Announce Type: new Abstract: Enterprise agents often need to integrate heterogeneous knowledge sources: documents for narrative facts, tables for computation, and dependency graphs

model-releasesarxiv-cs-cl
29 Jul 2026
Agents

A New Role for Relevance: Guiding Corpus Interaction in Agentic Search

DGX agent

arXiv:2607.24223v1 Announce Type: new Abstract: Relevance is a query-dependent estimate of whether a document or excerpt contains useful evidence. Existing retrieval agents use relevance to select top

agentsarxiv-cs-cl
28 Jul 2026
Model Releases

Accuracy Hides How Language Models Fail: Measuring Failure States Under Matched Output Budgets

DGX agent

arXiv:2607.24268v1 Announce Type: new Abstract: Language-model benchmarks collapse two distinct measurement questions into a single accuracy score: whether a response reached an evaluable state, and w

model-releasesarxiv-cs-cl
28 Jul 2026
Model Releases

Algorithmic Blindness in Large Language Models: A Calibration Study of Performance Prediction

DGX agent

arXiv:2602.21947v5 Announce Type: replace Abstract: Large language models (LLMs) demonstrate remarkable breadth of knowledge, yet their ability to reason about computational processes remains poorly u

model-releasesarxiv-cs-cl
28 Jul 2026
Research

An Efficient and Effective Evaluator for Text2SQL Models on Unseen and Unlabeled Data

DGX agent

arXiv:2603.07841v2 Announce Type: replace Abstract: Recent advances in large language models have strengthened Text2SQL systems that translate natural language questions into database queries. A persi

researcharxiv-cs-cl
28 Jul 2026
Model Releases

An MLIR-Based Compilation Method for Large Language Models

DGX agent

arXiv:2607.15865v2 Announce Type: replace Abstract: Large Language Models (LLMs) have become the dominant workload on modern AI accelerators, yet deploying them on specialized hardware still faces two

model-releasesarxiv-cs-cl
28 Jul 2026
Model Releases

AssumptionMiner: Extracting, Tracing, and Revising Implicit Assumptions in LLM Code Generation

DGX agent

arXiv:2607.22898v1 Announce Type: cross Abstract: Large language models (LLMs) generate code from natural-language prompts, yet real-world prompts rarely provide complete specifications. When prompts

model-releasesarxiv-cs-cl
28 Jul 2026
Model Releases

BERT-based Models vs. Large Language Models for Low-Resource Named Entity Recognition: A Comparative Study on Marathi

DGX agent

arXiv:2607.23344v1 Announce Type: new Abstract: Named Entity Recognition (NER) for low-resource languages such as Marathi remains a challenging task due to limited annotated resources and linguistic c

model-releasesarxiv-cs-cl
28 Jul 2026
Model Releases

Beyond Scale and Generation: Understanding Language Model-based Entity Matching

DGX agent

arXiv:2607.24688v1 Announce Type: cross Abstract: Entity matching identifies records that refer to the same real-world entity. Language models can be adapted to this task through bi-encoder, cross-enc

model-releasesarxiv-cs-cl
28 Jul 2026
Model Releases

BHARATI: Morphology-Aware Tokenizers for Classical Indian Languages with Subword Fertility Analysis

DGX agent

arXiv:2607.23319v1 Announce Type: new Abstract: Standard subword tokenization algorithms such as Byte-Pair Encoding (BPE) and SentencePiece are trained predominantly on modern language corpora and pro

model-releasesarxiv-cs-cl
28 Jul 2026
Research

Bigger or Cheaper? Scale and Quantization Effects on Uncertainty Signals in Vision-Language Models Under Image Degradation

DGX agent

arXiv:2607.24440v1 Announce Type: cross Abstract: Vision-language models (VLMs) deployed on consumer hardware must decide when to answer and when to defer, and that decision depends on having a confid

researcharxiv-cs-cl
28 Jul 2026
Model Releases

BioProBench: A Corpus and Benchmark for Biological Protocol Reasoning in Autonomous Science

DGX agent

arXiv:2505.07889v4 Announce Type: replace Abstract: The realization of autonomous scientific experimentation is currently limited by LLMs' struggle to grasp the strict procedural logic and accuracy re

model-releasesarxiv-cs-cl
28 Jul 2026
Research

BioSentinel at EXIST 2026: Soft-Label Optimization with XLM-RoBERTa for Sexism Intent Classification in Memes

DGX agent

arXiv:2607.24137v1 Announce Type: new Abstract: This paper describes the BioSentinel team's participation in EXIST 2026 Task 2.2: Source Intention in Memes, part of the CLEF 2026 evaluation campaign.

researcharxiv-cs-cl
28 Jul 2026
Research

CAGE: Cognitive Attribution Graphs for Faithful Inline Citation Generation in Long-Form Question Answering

DGX agent

arXiv:2607.24236v1 Announce Type: new Abstract: Long-form question answering increasingly relies on retrieved evidence to make LLM outputs verifiable, with inline citations tracing claims to source do

researcharxiv-cs-cl
28 Jul 2026
Model Releases

CausalGate: Causal Importance Distillation for Transformer Module Pruning

DGX agent

arXiv:2607.22720v1 Announce Type: cross Abstract: Existing adaptive inference methods for Large Language Models rely on observational heuristics, such as hidden-state similarity or activation magnitud

model-releasesarxiv-cs-cl
28 Jul 2026
Research

CHiPS: Character Histograms and Positional Signals for Lightweight Authorship Attribution in Romanian Texts

DGX agent

arXiv:2607.22884v1 Announce Type: new Abstract: We propose CHiPS, a lightweight character-level authorship attribution method for Romanian texts. All reported experiments are closed-set: the true auth

researcharxiv-cs-cl
28 Jul 2026
Research

Co-Evolving Graph and Text Memory for Training-Free Multi-Hop Question Answering

DGX agent

arXiv:2607.23278v1 Announce Type: new Abstract: Multi-hop question answering requires coordinating relational and textual evidence across reasoning steps, a combination neither a text corpus nor a kno

researcharxiv-cs-cl
28 Jul 2026
Local Ai

CONSISTRE: A Unified Consistency-Aware Framework for Document-Level Relation Extraction with Large Language Models

DGX agent

arXiv:2607.24312v1 Announce Type: new Abstract: Document-level relation extraction (DocRE) aims to extract relations among multiple entities across extended contexts while maintaining consistency acro

local-aiarxiv-cs-cl
28 Jul 2026
Research

Cross-Attention Calibrated Deduplication for Retrieval-Augmented Generation System

DGX agent

arXiv:2607.24332v1 Announce Type: new Abstract: Common chunking strategies in Retrieval-Augmented Generation (RAG) systems often create redundant chunks. These redundant chunks make the vector databas

researcharxiv-cs-cl
28 Jul 2026
Safety

Dependency-Guided Code Generation: Structured Matrix Decomposition and Consistency-Guided Refinement

DGX agent

arXiv:2607.16692v2 Announce Type: replace-cross Abstract: The increasing complexity of modern software systems has made automated code generation a fundamental task in software engineering. However, e

safetyarxiv-cs-cl
28 Jul 2026
Model Releases

Do Current Retrievers Cover All the Evidence? A Controlled Study of Conjunctive Cross-Page Retrieval

DGX agent

arXiv:2607.24165v1 Announce Type: cross Abstract: Finding a long document relevant to a multi-part request is not the same as establishing that it contains every requested piece of evidence. We study

model-releasesarxiv-cs-cl
28 Jul 2026
Safety

Do LLM Debates Repeat Arguments Differently Across Languages?

DGX agent

arXiv:2607.23442v1 Announce Type: new Abstract: LLM debate is usually evaluated by final answers, but transcripts also reveal whether later turns develop new argumentative content or return to earlier

safetyarxiv-cs-cl
28 Jul 2026
Safety

Does Faithfulness-Guided Alignment Hurt Accuracy? Unlocking Accurate and Faithful Post-Retrieval Reasoning

DGX agent

arXiv:2602.01348v3 Announce Type: replace Abstract: Retrieval-augmented generation (RAG) can achieve strong answer accuracy on multi-hop questions, but outcome-level rewards often leave reasoning trac

safetyarxiv-cs-cl
28 Jul 2026
Research

EmoTrace: An Emotion Trajectory-Centered Framework for Psychological Support Dialogue Generation

DGX agent

arXiv:2607.23648v1 Announce Type: new Abstract: Using large language models (LLMs) to assist psychological counseling is an important task in the field of natural language processing. The construction

researcharxiv-cs-cl
28 Jul 2026
Research

Ensembling LLM-Induced Decision Trees for Explainable and Robust Error Detection

DGX agent

arXiv:2512.07246v2 Announce Type: replace Abstract: Error detection (ED), which aims to identify incorrect or inconsistent cell values in tabular data, is important for ensuring data quality. Recent s

researcharxiv-cs-cl
28 Jul 2026
Research

Evidence Attribution in Visual Document Understanding without Coordinates or Region Labels

DGX agent

arXiv:2607.24651v1 Announce Type: cross Abstract: Reliable visual document understanding requires a model to attribute each answer to the evidence regions that support it. Recent benchmarks and system

researcharxiv-cs-cl
28 Jul 2026
Research

Explaining GAND: A Resource on Gender-Ambiguous Natural Data & Contrastive Attribution

DGX agent

arXiv:2607.22546v1 Announce Type: new Abstract: Machine translation (MT) systems continue to produce gender-biased translations. In a time where self-expression is paramount, mistranslations based on

researcharxiv-cs-cl
28 Jul 2026
Research

Formally Verified Synthesizable Floating-Point Data Types in ARCH HDL

DGX agent

arXiv:2607.23715v1 Announce Type: new Abstract: We report the design and end-to-end verification of first-class IEEE-754 binary32 (FP32) and bfloat16 (BF16) arithmetic for ARCH, a hardware description

researcharxiv-cs-cl
28 Jul 2026
Model Releases

From Data to Device: ELMOD An Efficient German-First 2.7B Language Model for Mobile Inference

DGX agent

arXiv:2607.24585v1 Announce Type: new Abstract: We present ELMOD - Efficient Language Model for On-Device Deployment - a compact (2.7B) German language model designed for efficient inference on resour

model-releasesarxiv-cs-cl
28 Jul 2026
Research

From peer review nuances to best practices

DGX agent

arXiv:2607.22681v1 Announce Type: cross Abstract: This report studies three nuances in peer review data: paper version, score version, and input format. We characterize how the variants differ, and me

researcharxiv-cs-cl
28 Jul 2026
← Previous
1…2122232425…161
Next →