AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlog
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
Model Releases

CORTEX: High-Quality Cross-Domain Organization of Web-Scale Corpora through Ontological Corpus Graph

DGX agent

arXiv:2606.30175v1 Announce Type: new Abstract: The continuous evolution of large language models drives escalating demands on data scale and quality, and as different training stages impose increasin

model-releasesarxiv-cs-cl
30 Jun 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Cross-Temporal Sinhala OCR: Page-Level Adaptation and Diachronic Analysis

DGX agent

arXiv:2606.29378v1 Announce Type: new Abstract: Sinhala is a morphologically rich abugida spoken by roughly 16 million people in Sri Lanka, and to date, there are no publicly available real-world data

model-releasesarxiv-cs-cl
30 Jun 2026
Agents

DAIN: Dynamic Agent-Based Interaction Network for Efficient and Collaborative Multimodal Reasoning

DGX agent

arXiv:2606.30189v1 Announce Type: new Abstract: Current multimodal fusion approaches, particularly those based on static Mixture-of-Experts (MoE) architectures, often struggle to provide the adaptive

agentsarxiv-cs-cl
30 Jun 2026
Model Releases

DataComp-VLM: Improved Open Datasets for Vision-Language Models

DGX agent

arXiv:2606.28551v1 Announce Type: cross Abstract: Building performant Vision-Language Models (VLMs) requires carefully curating large-scale training datasets, yet the community lacks systematic benchm

model-releasesarxiv-cs-cl
30 Jun 2026
Local Ai

Depth-Staggered Fibonacci Spacing for Sparse Attention: Static Schedules Beat Learned Dilation and Extrapolate Where Dense Attention Fails

DGX agent

arXiv:2606.28560v1 Announce Type: new Abstract: We study sparse self-attention in which each query attends to a dense local window plus a set of Fibonacci-spaced offsets, with a per-layer scalar alpha

local-aiarxiv-cs-cl
30 Jun 2026
Model Releases

Detecting Clinical Hallucinations in LVLMs via Counterfactual Visual Grounding Uncertainty

DGX agent

arXiv:2606.28520v1 Announce Type: cross Abstract: Large vision-language models (LVLMs) are increasingly used for clinical image understanding, yet they remain vulnerable to hallucinations--producing t

model-releasesarxiv-cs-cl
30 Jun 2026
Agents

Developmental Trajectories of Situation Modeling and Mentalizing in Transformer Language Models

DGX agent

arXiv:2606.28524v1 Announce Type: new Abstract: Recent work suggests that Large Language Models (LLMs) are sensitive to the belief states of agents described by text, as measured by the false belief t

agentsarxiv-cs-cl
30 Jun 2026
Model Releases

DIA-HARM: Dialectal Disparities in Harmful Content Detection Across 50 English Dialects

DGX agent

arXiv:2604.05318v2 Announce Type: replace Abstract: Harmful content detectors, particularly disinformation classifiers, are predominantly developed and evaluated on Standard American English (SAE), le

model-releasesarxiv-cs-cl
30 Jun 2026
Model Releases

DialogPII: A multilingual dataset of synthetic dialog transcripts to detect personal information

DGX agent

arXiv:2606.30312v1 Announce Type: new Abstract: Conversational data collected in domains such as healthcare or social sciences is a valuable resource for research and automated analysis. However, resp

model-releasesarxiv-cs-cl
30 Jun 2026
Model Releases

DistilledGemma: Balanced Efficiency-Accuracy for Person-Place Relation Extraction from Multilingual Historical Articles

DGX agent

arXiv:2606.29130v1 Announce Type: new Abstract: We present DistilledGemma, an efficient and accurate system for the HIPE-2026 shared task on person-place relation extraction from multilingual historic

model-releasesarxiv-cs-cl
30 Jun 2026
Model Releases

DNA Language Models: An Assessment of Pre-Training for Fine-Tuning Tasks

DGX agent

arXiv:2606.30140v1 Announce Type: cross Abstract: Recent breakthroughs in foundation models and Large Language Models (LLMs) have introduced new opportunities for studying and decoding genomic sequenc

model-releasesarxiv-cs-cl
30 Jun 2026
Safety

Do Models Read What They Write? Causal Registers in Scratchpad Reasoning

DGX agent

arXiv:2606.29522v1 Announce Type: cross Abstract: A central hope behind process supervision is that models can expose intermediate variables that matter for their later behavior. For this to help with

safetyarxiv-cs-cl
30 Jun 2026
Local Ai

DriftGuard: Safety-Aware Multi-Monitor Detection and Selective Adaptation for Evolving Toxicity Moderation

DGX agent

arXiv:2606.28725v1 Announce Type: new Abstract: Automated toxicity moderation systems operate in dynamic online environments where harmful behavior evolves through coded language, shifting targets, an

local-aiarxiv-cs-cl
30 Jun 2026
Research

Efficient Retrieval-Augmented Generation via Token Co-occurrence Graphs

DGX agent

arXiv:2606.30093v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) mitigates hallucinations in Large Language Models (LLMs) by grounding the generation process on external knowledge.

researcharxiv-cs-cl
30 Jun 2026
Safety

EntroRouter: Learning Efficient Model Routing via Entropy Regulation

DGX agent

arXiv:2606.29424v1 Announce Type: new Abstract: Model routing balances solution accuracy and computational cost by selecting among models of varying capabilities. While recent multi-round frameworks i

safetyarxiv-cs-cl
30 Jun 2026
Safety

EPIC-EuroParl-UdS: Information-Theoretic Perspectives on Translation and Interpreting

DGX agent

arXiv:2603.09785v3 Announce Type: replace Abstract: This paper introduces an updated and combined version of the bidirectional English-German EPIC-UdS (spoken) and EuroParl-UdS (written) corpora conta

safetyarxiv-cs-cl
30 Jun 2026
Research

Evaluating LLMs on Chinese Topic Constructions: A Research Proposal Inspired by Tian et al. (2024)

DGX agent

arXiv:2504.14969v2 Announce Type: replace Abstract: This paper proposes a framework for evaluating large language models (LLMs) on Chinese topic constructions, focusing on their sensitivity to island

researcharxiv-cs-cl
30 Jun 2026
Model Releases

EVLA: An Electro-Aware Multimodal Assistant for Physically-Grounded Driving Reasoning and Control

DGX agent

arXiv:2606.28938v1 Announce Type: new Abstract: Modern vision-language models (VLMs) for driving assistants typically treat vehicle dynamics as a black box, resulting in decisions that lack awareness

model-releasesarxiv-cs-cl
30 Jun 2026
Hardware

Evolution Fine-Tuning: Learning to Discover Across 371 Optimization Tasks

DGX agent

arXiv:2606.29082v1 Announce Type: new Abstract: Would experience designing faster GPU kernels also help close in on a long-standing open mathematical conjecture? Large Language Models (LLMs) integrate

hardwarearxiv-cs-cl
30 Jun 2026
Research

Exploiting Vision Encoder Vulnerabilities for Universal Adversarial Perturbations on Large Vision-Language Models

DGX agent

arXiv:2412.08108v3 Announce Type: replace-cross Abstract: Large Vision-Language Models (LVLMs) have achieved remarkable performance on multimodal tasks but remain highly vulnerable to small adversaria

researcharxiv-cs-cl
30 Jun 2026
Research

Extracting Knowledge from an Arabic-English Machine-Readable Dictionary Using Information Extraction

DGX agent

arXiv:2606.28457v1 Announce Type: new Abstract: Natural language processing (NLP) applications need large and rich amount of linguistic knowledge. Furthermore, electronic language sources such as dict

researcharxiv-cs-cl
30 Jun 2026
Research

Fast Numbers, Slow Language: Bridging Quantitative and Qualitative Earnings Signals

DGX agent

arXiv:2606.29734v1 Announce Type: new Abstract: Earnings announcements release two types of information sequentially: quantitative surprise (numeric earnings-per-share (EPS)/revenue versus analyst est

researcharxiv-cs-cl
30 Jun 2026
Research

FinInvest-GTCN: Explainable Graph-Temporal-Causal Modeling for Risk-Aware Investment Decision Optimization

DGX agent

arXiv:2606.28933v1 Announce Type: new Abstract: Venture capital (VC) investment decisions face distinct challenges, such as multi-source heterogeneous data, non-stationary time series, and the demand

researcharxiv-cs-cl
30 Jun 2026
Safety

Fund2Persona: A Framework for Building and Refining Financial Advisor Personas from Fund Disclosure Data

DGX agent

arXiv:2606.29793v1 Announce Type: new Abstract: Demand for personalized financial advising is growing, but consistent advisor expertise is difficult to obtain, scale, and encode in LLM systems. Simple

safetyarxiv-cs-cl
30 Jun 2026
Research

Generating in the Limit with Infinitely Many Hallucinations

DGX agent

arXiv:2606.28354v1 Announce Type: new Abstract: The classic paradigm of language identification in the limit models learning as a game between an adversary, who reveals strings from an unknown target

researcharxiv-cs-cl
30 Jun 2026
Research

Generative Large Language Models in Automated Fact-Checking: A Survey

DGX agent

arXiv:2407.02351v3 Announce Type: replace Abstract: The rapid spread of false and misleading information on online platforms poses a growing societal challenge, overwhelming the capacity of manual fac

researcharxiv-cs-cl
30 Jun 2026
Tutorials

Grounding LLM Reasoning under Incomplete Graph Evidence

DGX agent

arXiv:2606.30247v1 Announce Type: new Abstract: Knowledge graphs can guide large language models (LLMs) reasoning, but the graph seen by a system is usually a retrieved, linked, temporally scoped, and

tutorialsarxiv-cs-cl
30 Jun 2026
Model Releases

How Far Do On-Prem Open LLMs Get on Text-to-SQL? A Cross-Family Size x Technique Frontier on BIRD

DGX agent

arXiv:2606.29733v1 Announce Type: new Abstract: Organizations that cannot send data to a cloud API increasingly ask: how good is Text-to-SQL if the model must run on-premises on open weights, and whic

model-releasesarxiv-cs-cl
30 Jun 2026
Model Releases

How LLMs See Creativity: Zero-Shot Scoring of Visual Creativity with Interpretable Reasoning

DGX agent

arXiv:2606.29672v1 Announce Type: new Abstract: Evaluating the originality of visual images poses enduring challenges for creativity assessment. Automated scoring using AI models has proven effective

model-releasesarxiv-cs-cl
30 Jun 2026
Model Releases

How much of an LLM-generated clinical corpus is actually new? A production-scale measurement of content redundancy for provenance classification

DGX agent

arXiv:2606.29605v1 Announce Type: new Abstract: Clinical machine learning increasingly relies on training corpora generated by large language models (LLMs) rather than annotated by clinicians, and suc

model-releasesarxiv-cs-cl
30 Jun 2026
Safety

IHDec: Divergence-Steered Contrastive Decoding for Securing Multi-Turn Instruction Hierarchies

DGX agent

arXiv:2606.29960v1 Announce Type: new Abstract: Large Language Models (LLMs) often fail to maintain instruction hierarchies (IH) when processing multi-source inputs with varying role-level priorities,

safetyarxiv-cs-cl
30 Jun 2026
Research

Improving Large-Scale Weakly Supervised ASR by Filtering and Selection

DGX agent

arXiv:2606.28728v1 Announce Type: cross Abstract: Leveraging large-scale weakly supervised datasets is crucial to train robust end-to-end automatic speech recognition (ASR) models. However, such datas

researcharxiv-cs-cl
30 Jun 2026
Research

Information Dynamics of Language Communication

DGX agent

arXiv:2606.30096v1 Announce Type: new Abstract: Quantifying how meaning propagates through communicative exchanges remains underdeveloped in computational linguistics. Here we introduce an information

researcharxiv-cs-cl
30 Jun 2026
Agents

KbSD: Knowledge Boundary aware Self-Distillation for Behavioral Calibration in Agentic Search

DGX agent

arXiv:2606.29863v1 Announce Type: new Abstract: Agentic search equips large language models with dynamic retrieval abilities, but existing reinforcement learning methods remain limited by reward spars

agentsarxiv-cs-cl
30 Jun 2026
Model Releases

Know Before You Fetch: Calibrated Retrieval-Budget Allocation for Retrieval-Augmented Generation

DGX agent

arXiv:2606.29959v1 Announce Type: cross Abstract: Retrieval-augmented generation (RAG) typically retrieves a fixed number of passages for every query. This is wasteful when the reader already knows th

model-releasesarxiv-cs-cl
30 Jun 2026
Model Releases

Labeling Training Data for Entity Matching Using Large Language Models

DGX agent

arXiv:2606.28823v1 Announce Type: new Abstract: Recent large language models (LLMs) achieve strong performance on entity matching without requiring task-specific training data. However, applying these

model-releasesarxiv-cs-cl
30 Jun 2026
Safety

LatentRevise: Learning from Zero-Hit Reasoning

DGX agent

arXiv:2606.29938v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) is bottlenecked by hard prompts on which correct trajectories have low probability, so sampling mi

safetyarxiv-cs-cl
30 Jun 2026
Applications

Legal Domain Adaptation of Modern BERT Models

DGX agent

arXiv:2606.28538v1 Announce Type: new Abstract: We investigate domain adaptation of modern BERT models in the legal domain. We further pre-train ModernBERT on all US court opinions using the masked la

applicationsarxiv-cs-cl
30 Jun 2026
Model Releases

LLM Agents Are Latent Context Managers: Eliciting Self-Managed Context via a Proprioceptive Dashboard

DGX agent

arXiv:2606.30005v1 Announce Type: new Abstract: Long-horizon tool agents are bottlenecked by how their context grows toward the limits of the context window. Recent systems make context management age

model-releasesarxiv-cs-cl
30 Jun 2026
Research

LLMs and their Limited Theory of Mind: Evaluating Mental State Annotations in Situated Dialogue

DGX agent

arXiv:2509.02292v2 Announce Type: replace Abstract: What if large language models could not only infer human mindsets but also expose every blind spot in team dialogue such as discrepancies in the tea

researcharxiv-cs-cl
30 Jun 2026
Model Releases

MaDI-Bench: An End-to-End Data Integration Benchmark

DGX agent

arXiv:2606.30371v1 Announce Type: cross Abstract: Data integration combines heterogeneous data sets into a single, coherent representation. Data integration involves a sequence of interdependent tasks

model-releasesarxiv-cs-cl
30 Jun 2026
Model Releases

MAM-AI: An On-Device Medical Retrieval-Augmented Generation System for Nurses and Midwives in Zanzibar

DGX agent

arXiv:2606.29580v1 Announce Type: new Abstract: Maternal and newborn mortality remain among the highest in sub-Saharan Africa, where midwifery care is often delivered by nurses who lack midwifery trai

model-releasesarxiv-cs-cl
30 Jun 2026
Model Releases

mamabench and mamaretrieval: Benchmarks for Evaluating Medical Retrieval-Augmented Generation in Maternal, Neonatal, and Reproductive Health

DGX agent

arXiv:2606.29467v1 Announce Type: new Abstract: Medical question-answering benchmarks rarely cover the maternal, neonatal, child, and reproductive-health questions a nurse-midwife asks, and, to our kn

model-releasesarxiv-cs-cl
30 Jun 2026
Research

Managing Map Cardinality in Automatic Disease Classification Mapping: Balancing Precision, Recall and Coverage

DGX agent

arXiv:2606.29750v1 Announce Type: new Abstract: Automatic mapping between disease classification systems, such as the International Classification of Diseases (ICD), is a challenging yet essential tas

researcharxiv-cs-cl
30 Jun 2026
Safety

Masked Diffusion Decoding as x-Prediction Flow

DGX agent

arXiv:2606.29066v1 Announce Type: new Abstract: Masked diffusion language models (MDLMs) generate text by iteratively unmasking tokens, but their standard decoder reduces each step to a binary action:

safetyarxiv-cs-cl
30 Jun 2026
Tutorials

MauBERT: Universal Phonetic Inductive Biases for Few-Shot Acoustic Units Discovery

DGX agent

arXiv:2512.19612v2 Announce Type: replace Abstract: This paper introduces MauBERT, a multilingual extension of HuBERT that leverages articulatory features for robust cross-lingual phonetic representat

tutorialsarxiv-cs-cl
30 Jun 2026
Model Releases

MemDelta: Controlled Baselines and Hidden Confounds in Agent Memory Evaluation

DGX agent

arXiv:2606.29914v1 Announce Type: new Abstract: Agent memory systems are increasingly evaluated against RAG and full-context baselines, but reported gains often mix changes in the memory method with c

model-releasesarxiv-cs-cl
30 Jun 2026
Model Releases

Memory-Managed Long-Context Attention: A Preliminary Study of Editable Request-Local Memory

DGX agent

arXiv:2606.28876v1 Announce Type: new Abstract: Long-context language models often conflate two different goals: compressing history into an efficient state, and maintaining reliable long-term memory.

model-releasesarxiv-cs-cl
30 Jun 2026
← Previous
1…3738394041…161
Next →