AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlog
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
Agents

To Isolate or to Score? Model-Adaptive Assessment for Cost-Efficient Multi-Agent RAG

DGX agent

arXiv:2606.25191v1 Announce Type: cross Abstract: Multi-agent document assessment for retrieval-augmented generation is computationally expensive, driving practitioners toward smaller, deployable mode

agentsarxiv-cs-cl
25 Jun 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Toten: A Knowledge-Based System For Structure-Preserving Representation Of Physical Quantities And Technical Notation In Brazilian Portuguese

DGX agent

arXiv:2606.19626v2 Announce Type: replace-cross Abstract: AI pipelines that reason quantitatively over technical text depend on input where physical quantities, numbers, units, and symbolic expression

model-releasesarxiv-cs-cl
25 Jun 2026
Research

Towards Structuring an Arabic-English Machine-Readable Dictionary Using Parsing Expression Grammars

DGX agent

arXiv:2606.25231v1 Announce Type: new Abstract: Dictionaries are rich sources of lexical information about words that is required for many applications of natural language processing and human languag

researcharxiv-cs-cl
25 Jun 2026
Research

Tracing Target Answers in Poisoned Retrieval Corpora via Token Influence Attribution

DGX agent

arXiv:2606.25721v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) systems are vulnerable to corpus poisoning attacks that manipulate model outputs through malicious retrieved docu

researcharxiv-cs-cl
25 Jun 2026
Model Releases

Uncertainty Quantification for Computer-Use Agents: A Benchmark across Vision-Language Models and GUI Grounding Datasets

DGX agent

arXiv:2606.25760v1 Announce Type: cross Abstract: Computer-use agents turn vision-language model (VLM) predictions into executable GUI clicks, so reliable uncertainty estimates are essential for rejec

model-releasesarxiv-cs-cl
25 Jun 2026
Agents

VADAOrchestra: Neurosymbolic Orchestration of Adaptive Reasoning Workflows

DGX agent

arXiv:2606.22485v2 Announce Type: replace-cross Abstract: Decision-making in real-world settings rarely follows a fixed script. Instead, it unfolds as a dynamic reasoning process in which the appropri

agentsarxiv-cs-cl
25 Jun 2026
Research

Weave of Formal Thought

DGX agent

arXiv:2606.25987v1 Announce Type: new Abstract: Large language models (LLMs) attain remarkable surface fluency on code, yet they neither formally guarantee the syntactic validity of their output nor l

researcharxiv-cs-cl
25 Jun 2026
Model Releases

What Intermediate Layers Know: Detecting Jailbreaks from Entropy Dynamics

DGX agent

arXiv:2606.25182v1 Announce Type: new Abstract: Jailbreak attacks reveal a persistent weakness in aligned Large Language Models: carefully crafted prompts can elicit policy-violating responses despite

model-releasesarxiv-cs-cl
25 Jun 2026
Research

When Certainty Is an Artifact: Keyword Lexicon Blindness and the (Mis)Measurement of Rhetorical Stance

DGX agent

arXiv:2606.26062v1 Announce Type: new Abstract: Can a statistically significant, large-effect-size finding in computational social science be entirely an artifact of the measurement instrument? We pre

researcharxiv-cs-cl
25 Jun 2026
Research

Why Do Accumulated Transformations Extrapolate?

DGX agent

arXiv:2606.24975v1 Announce Type: cross Abstract: PaTH Attention showed that replacing RoPE's position-indexed rotations with accumulated data-dependent Householder reflections yields strong length ex

researcharxiv-cs-cl
25 Jun 2026
Safety

Why Multi-Step Tool-Use Reinforcement Learning Collapses and How Supervisory Signals Fix It

DGX agent

arXiv:2606.26027v1 Announce Type: new Abstract: Tool use enables large language models (LLMs) to perform complex tasks, and recent agentic reinforcement learning (RL) methods show promise for enhancin

safetyarxiv-cs-cl
25 Jun 2026
Model Releases

A Synthetic Reliability-Aware PINN Benchmark for Offshore Wind Turbine Support-Structure Monitoring with Bayesian Inverse Identification

DGX agent

arXiv:2606.24176v1 Announce Type: new Abstract: Reliable structural health monitoring (SHM) of offshore wind turbine (OWT) support structures requires fast state estimation from sparse measurements. R

model-releasesarxiv-cs-cl
24 Jun 2026
Model Releases

AGORA: An Archive-Grounded Benchmark for Agentic Workplace Document Reasoning

DGX agent

arXiv:2606.24526v1 Announce Type: new Abstract: Large language models are increasingly deployed as agents that reason over documents rather than answer from parametric knowledge. We study archive-grou

model-releasesarxiv-cs-cl
24 Jun 2026
Safety

An LLM-based Two-Stage Transformer Framework for Cross-Domain Bearing Fault Diagnosis with Limited Data

DGX agent

arXiv:2606.24459v1 Announce Type: cross Abstract: Bearing fault diagnosis faces critical challenges when dataset heterogeneity, operating condition variations, and limited labeled data occur simultane

safetyarxiv-cs-cl
24 Jun 2026
Model Releases

Are We Ready For An Agent-Native Memory System?

DGX agent

arXiv:2606.24775v1 Announce Type: new Abstract: Memory for large language model (LLM) agents has rapidly evolved from simple retrieval-augmented mechanisms into a data management system that supports

model-releasesarxiv-cs-cl
24 Jun 2026
Research

Aspect-Based Sentiment Evolution and its Correlation with Review Rounds in Multi-Round Peer Reviews: A Deep Learning Approach

DGX agent

arXiv:2606.24188v1 Announce Type: new Abstract: Mining sentiment information from the textual content of peer review comments offers valuable insights into the scientific evaluation process. However,

researcharxiv-cs-cl
24 Jun 2026
Research

Automatic Part-of-Speech Tagging of Arabic-English Dictionary Senses through WordNet

DGX agent

arXiv:2606.24359v1 Announce Type: new Abstract: This paper proposed an algorithm for part-of-speech (POS) tagging senses of a bilingual dictionary. The algorithm is applied on the Al-Mawrid Arabic-Eng

researcharxiv-cs-cl
24 Jun 2026
Model Releases

AutoSpecNER: A Fine-Grained Named Entity Recognition Dataset for Vehicle Specification Extraction

DGX agent

arXiv:2606.24387v1 Announce Type: new Abstract: Vehicle advertisements contain rich specification information, but automotive NER resources remain limited. We introduce AutoSpecNER, an expert-annotate

model-releasesarxiv-cs-cl
24 Jun 2026
Research

AVOC: Enhancing Hour-Level Audio-Video Understanding in Omni-Modal LLMs via Retrieval-Inspired Token Compression

DGX agent

arXiv:2606.24286v1 Announce Type: new Abstract: Multimodal Large Language Models have achieved remarkable progress in short-form audio-video understanding, yet long-form audio-video comprehension rema

researcharxiv-cs-cl
24 Jun 2026
Model Releases

BehaviorBench: Benchmarking Foundation Models for Behavioral Science Tasks

DGX agent

arXiv:2606.24162v1 Announce Type: new Abstract: Foundation models have been increasingly applied to behavioral science domains such as psychology, sociology, and economics. While these models show pro

model-releasesarxiv-cs-cl
24 Jun 2026
Applications

Best Preprocessing Techniques for Sentiment Analysis

DGX agent

arXiv:2606.24055v1 Announce Type: new Abstract: Sentiment analysis in Twitter datasets is important because it enables monitoring public opinion on products and analysis of political and social moveme

applicationsarxiv-cs-cl
24 Jun 2026
Research

Beyond Logprobs: A Multi-Signal Confidence Engine for LLM-Based Document Field Extraction

DGX agent

arXiv:2606.24420v1 Announce Type: new Abstract: In high-stakes document processing pipelines, including financial reconciliation, compliance verification, and procurement automation, an LLM extraction

researcharxiv-cs-cl
24 Jun 2026
Safety

Bilevel Data Curation for LLM Fine-tuning: Offline Selection and Online Self-Refining Generation

DGX agent

arXiv:2511.21056v2 Announce Type: replace-cross Abstract: Supervised fine-tuning (SFT) datasets are critical to the downstream performance of large language models, yet they often contain low-quality

safetyarxiv-cs-cl
24 Jun 2026
Model Releases

Business as Rulesual: A Benchmark and Framework for Business Rule Flow Modeling with LLMs

DGX agent

arXiv:2505.18542v4 Announce Type: replace Abstract: Extracting structured procedural knowledge from unstructured business documents is a critical yet unresolved bottleneck in process automation. While

model-releasesarxiv-cs-cl
24 Jun 2026
Model Releases

CANDLE: Character-level Arabic Noise Deduplication using Lightweight Encoder

DGX agent

arXiv:2606.24758v1 Announce Type: new Abstract: Handling repeated characters in text can be tricky, since they can represent either the correct spelling of a word or informal character elongation ofte

model-releasesarxiv-cs-cl
24 Jun 2026
Model Releases

CN-NewsTTS Bench: a target-level automatic benchmark for raw-input Chinese news TTS pronunciation

DGX agent

arXiv:2606.24714v1 Announce Type: new Abstract: Chinese news text contains dense written forms such as scores, hyphenated model names, ranges, unit symbols, percentages, English abbreviations, and mix

model-releasesarxiv-cs-cl
24 Jun 2026
Research

ComputeFHE: A Privacy-Preserving General-Purpose Computation Library

DGX agent

arXiv:2606.24379v1 Announce Type: cross Abstract: Fully Homomorphic Encryption (FHE) enables computations to be performed directly on encrypted data while preserving data confidentiality. However, its

researcharxiv-cs-cl
24 Jun 2026
Research

CORE-BREW: LLR-Based Soft Decoding for Robust Multi-Bit LLM Watermarking

DGX agent

arXiv:2606.24163v1 Announce Type: cross Abstract: Reliable provenance for LLM outputs requires multi-bit watermarks that remain robust under editing while maintaining strict false-positive control. Ex

researcharxiv-cs-cl
24 Jun 2026
Local Ai

Cross-Lingual Exploration for Parametric Knowledge

DGX agent

arXiv:2606.24579v1 Announce Type: new Abstract: Parametric knowledge in Large Language Models is not equally accessible across languages. As a result, standard inference techniques often struggle to s

local-aiarxiv-cs-cl
24 Jun 2026
Local Ai

Decoherence as Defence and the Magnitude of Noise Regularisation: A Rigorous N -Qubit Theory of Stochastic Quantum Neural Networks for Adversarially Robust Network Intrusion Detection

DGX agent

arXiv:2606.24219v1 Announce Type: new Abstract: Stochastic quantum neural networks (SQNNs) encode neuronal activations as qubits, synaptic topology as entanglement, and neural noise through a Lindblad

local-aiarxiv-cs-cl
24 Jun 2026
Research

Dialogue to Discovery: Attribute-Aware Preference Elicitation for Conversational Product Search Assistants

DGX agent

arXiv:2606.24194v1 Announce Type: cross Abstract: Conversational product search assistants offer a more expressive, natural, and interactive alternative to traditional keyword-based product search. Wi

researcharxiv-cs-cl
24 Jun 2026
Research

Do LLM Attribution Metrics Transfer? Auditing Retrieval-Augmented Generation Evaluation Across Datasets and Constructs

DGX agent

arXiv:2606.23915v1 Announce Type: new Abstract: Practice often treats automatic metrics for attribution in LLM retrieval-augmented generation as interchangeable. We audit eight automatic scorers -- le

researcharxiv-cs-cl
24 Jun 2026
Research

Does My Embedding Reflect That A = B? Evaluating Mathematical Equivalence in Embedding Models

DGX agent

arXiv:2606.23959v1 Announce Type: new Abstract: Because mathematics is highly abstract, a single statement can take very different forms depending on what subfield it is framed in. There are many exam

researcharxiv-cs-cl
24 Jun 2026
Research

DREAM: Dense Retrieval Embeddings via Autoregressive Modeling

DGX agent

arXiv:2606.24667v1 Announce Type: new Abstract: Dense retrieval embedding models are a fundamental component of modern retrieval-based AI systems. Most dense retrievers are trained with contrastive ob

researcharxiv-cs-cl
24 Jun 2026
Research

ErrorLLM: Modeling SQL Errors for Text-to-SQL Refinement

DGX agent

arXiv:2603.03742v2 Announce Type: replace Abstract: Despite the remarkable performance of large language models (LLMs) in text-to-SQL (SQL generation), correctly producing SQL queries remains challeng

researcharxiv-cs-cl
24 Jun 2026
Local Ai

ESBMC-GraphPLC: Formal Verification of Graphical PLCopen XML Ladder Diagram Programs Using SMT-Based Model Checking

DGX agent

arXiv:2606.18941v3 Announce Type: replace-cross Abstract: PLCopen XML defines two encoding formats for IEC 61131-3 Ladder Diagram programs: a textual encoding using elements, and a graphical encoding

local-aiarxiv-cs-cl
24 Jun 2026
Model Releases

ESBMC-PLC+: A Unified IEC~61131-3 Formal Verification Framework as a PLCverif Successor

DGX agent

arXiv:2606.23870v1 Announce Type: cross Abstract: PLCverif is the most mature open-source platform for PLC formal verification, developed at CERN and in production use since 2019. Yet it has two funda

model-releasesarxiv-cs-cl
24 Jun 2026
Safety

Escaping the Self-Confirmation Trap: An Execute-Distill-Verify Paradigm for Agentic Experience Learning

DGX agent

arXiv:2606.24428v1 Announce Type: new Abstract: Experience-driven self-evolution is critical for large language model (LLM) agents to improve through open-world interaction. However, existing experien

safetyarxiv-cs-cl
24 Jun 2026
Safety

EvidenceLens: A Claim-Evidence Matrix for Auditing Financial Question Answering

DGX agent

arXiv:2606.23724v1 Announce Type: cross Abstract: Large language models are increasingly used to answer questions over annual reports, earnings decks, and analyst notes, yet their outputs remain diffi

safetyarxiv-cs-cl
24 Jun 2026
Safety

EXPO-SQL: Execution-based Clause-level Policy Optimization for Text-to-SQL

DGX agent

arXiv:2606.23693v1 Announce Type: new Abstract: Text-to-SQL enables users to query databases using natural language by generating executable SQL queries. Recent methods have increasingly adopted Large

safetyarxiv-cs-cl
24 Jun 2026
Research

Few shot chain-of-thought driven reasoning to prompt LLMs for open ended medical question answering

DGX agent

arXiv:2403.04890v4 Announce Type: replace Abstract: In this paper, we propose a modified version of the MedQA-USMLE dataset, named MEDQA-OPEN, which contains open-ended medical questions without optio

researcharxiv-cs-cl
24 Jun 2026
Research

Ground Then Rank: Revisiting Knowledge-Based VQA with Training-Free Entity Identification

DGX agent

arXiv:2606.23881v1 Announce Type: new Abstract: Knowledge-Based Visual Question Answering (KB-VQA) requires grounding visual queries to external knowledge beyond directly observable content in images.

researcharxiv-cs-cl
24 Jun 2026
Model Releases

Harmonic: Hierarchical State Space Models for Efficient Long-Context Language Modeling

DGX agent

arXiv:2606.24650v1 Announce Type: new Abstract: We present Harmonic, a hierarchical state space model (SSM) for language modeling. The architecture stacks three recurrent levels at progressively slowe

model-releasesarxiv-cs-cl
24 Jun 2026
Model Releases

Holistic Data Scheduler for LLM Pre-training via Multi-Objective Reinforcement Learning

DGX agent

arXiv:2606.24133v1 Announce Type: cross Abstract: The composition of training data, governed by the diversity of sources and their mixing strategy, is a cornerstone of Large Language Model (LLM) pre-t

model-releasesarxiv-cs-cl
24 Jun 2026
Model Releases

How Much Can We Trust LLM Search Agents? Measuring Endorsement Vulnerability to Web Content Manipulation

DGX agent

arXiv:2606.16821v2 Announce Type: replace Abstract: Large language model (LLM)-based search agents synthesize open-web content into actionable recommendations on behalf of users, creating a risk that

model-releasesarxiv-cs-cl
24 Jun 2026
Model Releases

Knowledge-Graph Grounding Helps LLMs Only for Out-of-Training Knowledge: A Controlled Study on Clinical Question Answering

DGX agent

arXiv:2606.22419v2 Announce Type: replace Abstract: A recent Nature Medicine study reports that general-purpose frontier LLMs outperform specialized retrieval-augmented clinical tools on medical bench

model-releasesarxiv-cs-cl
24 Jun 2026
Model Releases

L3Cube-MahaPOS: A Marathi Part-of-Speech Tagging Dataset and BERT Models

DGX agent

arXiv:2606.24825v1 Announce Type: new Abstract: Part-of-Speech (POS) tagging is a foundational NLP task underpinning machine translation, information extraction, and syntactic parsing. Despite Marathi

model-releasesarxiv-cs-cl
24 Jun 2026
Model Releases

LangMAP: A Language-Adaptive Approach to Tokenization

DGX agent

arXiv:2606.23566v2 Announce Type: replace Abstract: Language-specific tokenizers improve tokenization quality and the downstream performance of models on those languages. However, using such a tokeniz

model-releasesarxiv-cs-cl
24 Jun 2026
← Previous
1…4445464748…161
Next →