AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
15 Jul 2026

TAKE: Trajectory-Aware Knowledge Estimation for Text Dataset Distillation

ResearchDGX agent

arXiv:2607.11898v1 Announce Type: new Abstract: Large-scale text corpora have become a quiet bottleneck in modern NLP, not just in storage, but in the accumulated cost of training, fine-tuning, and co

The Capacity of Thought: Benchmarking Llama 3.2 in Semantic fMRI Neural Language Decoding and Improving the Huth Encoding-Model Baseline

Model ReleasesDGX agent

arXiv:2607.12079v1 Announce Type: new Abstract: Decoding continuous language from fMRI signals remains a core challenge in non-invasive brain-computer interface research. We present two complementary

The Illusion of Robustness: Aggregate Accuracy Hides Prediction Flips under Task-Irrelevant Context

Model Releases

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2607.12963v1 Announce Type: new Abstract: As large language models (LLMs) grow more capable, they are increasingly deployed in context-rich settings where task inputs are often accompanied by lo

Token Reduction Is Not Cost Reduction

Model ReleasesDGX agent

arXiv:2607.12161v1 Announce Type: new Abstract: Context-reduction layers for API-based coding agents, including command-output compressors, retrieval rankers, and payload-optimizing proxies, are usual

Transforming LLMs into Efficient Cross-Encoders via Knowledge Distillation for RAG Reranking

Model ReleasesDGX agent

arXiv:2607.11933v1 Announce Type: new Abstract: Cross-encoders achieve high reranking accuracy in Retrieval-Augmented Generation (RAG) pipelines but impose quadratic inference costs that limit real-ti

Translation as a Computationally Efficient Bridge: Feasibility of English BERT for Low-Resource Languages

ResearchDGX agent

arXiv:2607.12612v1 Announce Type: new Abstract: BERT models have revolutionised Natural Language Processing (NLP) through their ability to process unstructured text across diverse domains. However, de

We Hebben Een Serieus Translatie: Modeling Intercomprehension as Probabilistic Inference

SafetyDGX agent

arXiv:2607.12169v1 Announce Type: new Abstract: Intercomprehension refers to partial intelligibility of an unfamiliar language (L2) by a speaker of a related language (L1). How is this zero-shot cross

WikiSTAR: A System for Shedding Light on the Hidden History of Scientific Wikipedia Articles

Model ReleasesDGX agent

arXiv:2607.12441v1 Announce Type: new Abstract: Wikipedia plays a key role in shaping public understanding of science, and its openly accessible revision history is a unique record of how scientific k

10 Jul 2026

COALA: Robust Contextualized Speech-augmented Language Modeling for ASR via Contrastive Regularizer and Biasing Score Estimation

Model ReleasesDGX agent

arXiv:2607.08117v1 Announce Type: new Abstract: Contextual biasing seeks to integrate external knowledge into automatic speech recognition (ASR) systems to accurately recognize domain-specific entitie

Cross-seed explainability using Procrustes-conditioned Joint End-to-end Top-K Sparse Autoencoders

Model ReleasesDGX agent

arXiv:2607.08499v1 Announce Type: new Abstract: We present a Procrustes-conditioned Joint End-to-end Top-K Sparse Autoencoder (SAE) for extracting cross-seed universal features from independently trai

DeepSearch-World: Self-Distillation for Deep Search Agents in a Verifiable Environment

AgentsDGX agent

arXiv:2607.07820v1 Announce Type: new Abstract: Training tool-use agents to improve from their own experience remains challenging, as supervised fine-tuning relies on fixed teacher-distilled trajector

Detecting Ladder Logic Bombs in IEC 61131-3 PLC Programs using ESBMC-PLC+: A Formal Verification Approach with Trigger Synthesis

SafetyDGX agent

arXiv:2607.08417v1 Announce Type: new Abstract: A Ladder Logic Bomb (LLB) is malicious control logic in a Programmable Logic Controller (PLC) program that lies dormant until a trigger activates a payl

Diarization-Guided Qwen-ASR Adaptation for Multilingual Two-Speaker Conversational Speech

Model ReleasesDGX agent

arXiv:2607.08208v1 Announce Type: new Abstract: This paper describes our self-designed system for Task 1 of the MLC-SLM 2026 Challenge for multilingual two-speaker conversational speech. The system co

Do You Need a Frontier Model as a Citation Verifier? Benchmarking Rubric LLMs for Deep-Research Source Attribution

Model ReleasesDGX agent

arXiv:2607.08700v1 Announce Type: new Abstract: Reinforcement learning increasingly relies on an LLM judge to score each rubric criterion, and that judge acts as the reward model during training. Befo

DominoTree: Conditional Tree-Structured Drafting with Domino for Speculative Decoding

Model ReleasesDGX agent

arXiv:2607.08642v1 Announce Type: new Abstract: Speculative decoding accelerates LLM inference by drafting several tokens and verifying them in parallel. Block-diffusion drafters such as DFlash produc

DR-Arena: an Automated Evaluation Framework for Deep Research Agents

SafetyDGX agent

arXiv:2601.10504v2 Announce Type: replace Abstract: As Large Language Models (LLMs) increasingly operate as Deep Research (DR) Agents capable of autonomous investigation and information synthesis, rel

Echoes Across Vietnam's Highlands, Delta, and Coast: A Multilingual Corpus for Cham, Khmer, and Tay-Nung

Model ReleasesDGX agent

arXiv:2607.08362v1 Announce Type: new Abstract: Vietnam's ethnic minority languages are almost absent from the field of Natural Language Processing (NLP), and the challenge goes beyond data scarcity:

Ensemble Diversity Optimization for Subjective Supervision

SafetyDGX agent

arXiv:2607.08493v1 Announce Type: cross Abstract: Subjective NLP tasks often exhibit systematic annotator disagreement, requiring models that represent uncertainty rather than collapse it. We introduc

Fair Document Valuation in LLM Summaries via Shapley Values

ResearchDGX agent

arXiv:2505.23842v5 Announce Type: replace Abstract: Large Language Models (LLMs) increasingly power search engines and AI assistants that retrieve and summarize content from many sources. By serving a

fog: Expressing Motion and Emotion through Function Composition of AI-Generated Code

ResearchDGX agent

arXiv:2607.07952v1 Announce Type: cross Abstract: Motion and emotion are core parts of intelligent, expressive behavior. In this paper, we introduce fog, a function composition framework for implement

From Execution to Education: A Bloom-Aligned Framework for Measuring Educational Control in LLMs

ApplicationsDGX agent

arXiv:2607.08009v1 Announce Type: new Abstract: We introduce a Bloom-aligned framework for measuring educational control in Large Language Models (LLMs): the ability to preserve a task's instructional

Grounded Event Extraction from SEC 8-K Filings with a Fine-Grained Taxonomy

ResearchDGX agent

arXiv:2607.08346v1 Announce Type: new Abstract: Form 8-K filings are the primary channel through which U.S. public companies disclose material events, but the SEC item codes attached to them are coars

Hallucination Self-Play: Bootstrapping Reinforced Detector via Evolved Generator

Model ReleasesDGX agent

arXiv:2607.07993v1 Announce Type: new Abstract: Identifying faithfulness hallucinations in LLM-generated outputs remains challenging due to the scarcity of high-quality annotated data. Recent work rel

HeaPA: Difficulty-Aware Heap Sampling and On-Policy Query Augmentation for LLM Reinforcement Learning

SafetyDGX agent

arXiv:2601.22448v2 Announce Type: replace-cross Abstract: RLVR has become a standard recipe for training LLMs on reasoning tasks with verifiable outcomes, but when rollout generation dominates the cos

Hidden Decoding at Scale: Latent Computation Scaling for Large Language Models

ResearchDGX agent

arXiv:2607.08186v1 Announce Type: new Abstract: Scaling Large Language Models (LLMs) has been driven mainly by enlarging the Transformer backbone, but for an already-strong model this requires another

Holographic Neural PCFG for Unsupervised Parsing

ResearchDGX agent

arXiv:2607.08063v1 Announce Type: new Abstract: Unsupervised constituency parsing aims to accurately induce latent tree structures from raw text alone. Recent neural parameterizations of PCFGs achieve

IMProofBench: Benchmarking AI on Research-Level Mathematical Proof Generation

Model ReleasesDGX agent

arXiv:2509.26076v2 Announce Type: replace Abstract: As the mathematical capabilities of large language models (LLMs) improve, it becomes increasingly important to evaluate their performance on researc

Improving Ad-hoc Search Effectiveness for Conversational Information Retrieval via Model Merging

Model ReleasesDGX agent

arXiv:2607.08540v1 Announce Type: cross Abstract: Conversational information retrieval is challenging since it requires the consideration of the conversation history which potentially gives rise to to

It Takes a MAESTRO To Prune Bad Experts

SafetyDGX agent

arXiv:2607.08601v1 Announce Type: new Abstract: Sparsely-activated Mixture-of-Experts (MoE) language models achieve remarkable inference efficiency by activating only a small fraction of parameters pe

MASTE: A Multi-Agent Pipeline for Zero-Shot Aspect Sentiment Triplet Extraction

AgentsDGX agent

arXiv:2607.08080v1 Announce Type: new Abstract: Aspect Sentiment Triplet Extraction (ASTE) requires jointly identifying (aspect, opinion, sentiment) triples from a given review sentence. While large l

Prompt Compression via Activation Aggregation

ResearchDGX agent

arXiv:2607.08399v1 Announce Type: new Abstract: Large language models process prompts by propagating activations through dozens of layers before generating a response. We ask whether the task-relevant

Scalable and Culturally Specific Stereotype Dataset Construction via Human-LLM Collaboration

SafetyDGX agent

arXiv:2607.07895v1 Announce Type: new Abstract: Research on stereotypes in large language models (LLMs) has largely focused on English-speaking contexts, due to the lack of datasets in other languages

Semantic Representation Learning of Scientific Literature based on Adaptive Feature and Graph Neural Network

ResearchDGX agent

arXiv:2311.00296v2 Announce Type: replace Abstract: Because most scientific literature data are unlabeled, semantic representation learning based on unsupervised graphs has become crucial. To enrich s

SPL: Orchestrating Workflows with Declarative Deterministic-Probabilistic Composition

Local AiDGX agent

arXiv:2607.07727v1 Announce Type: cross Abstract: We present SPL (Structured Prompt Language), a declarative language that composes deterministic and probabilistic computation modes in a single specif

SQuaD-SQL: Efficient Text-to-SQL with Small Language Models via LLM-Guided Knowledge Distillation

Model ReleasesDGX agent

arXiv:2607.08161v1 Announce Type: new Abstract: Text-to-SQL is a fundamental task in natural language processing that enables users to interact with structured databases using natural language. While

The Memory Wall of Green Software: Empirical Energy Evaluation of Memento Design Pattern

ResearchDGX agent

arXiv:2607.07944v1 Announce Type: cross Abstract: As Green Software Engineering matures, energy efficiency has transitioned into a mission-critical non-functional requirement. While software design pa

Token-Flow Firewall: Semantic Runtime Auditing for Persistent AI Agents

Local AiDGX agent

arXiv:2607.08395v1 Announce Type: cross Abstract: Persistent AI agents extend large language models (LLMs) beyond single-turn interaction into long-lived software systems. Unlike traditional chat assi

Tool-Making and Self-Evolving LLM Agents in Low-Latency Systems

AgentsDGX agent

arXiv:2607.08010v1 Announce Type: new Abstract: Production LLM agents often waste latency and reliability by regenerating code for the same procedural steps on every request. We replace this inference

Truncated Step-Level Sampling with Process Rewards for Retrieval-Augmented Reasoning

ResearchDGX agent

arXiv:2602.23440v4 Announce Type: replace Abstract: Reinforcement learning has emerged as an effective paradigm for training large language models to interleave reasoning with search engine calls. How

Uncertainty-gated selection for block-sparse attention

Model ReleasesDGX agent

arXiv:2607.07724v1 Announce Type: cross Abstract: Block-sparse attention scales long-context language models by replacing the O(N^2) softmax with a per-query top-k selection over key blocks. This cuto

UniClawBench: A Universal Benchmark for Proactive Agents on Real-World Tasks

Model ReleasesDGX agent

arXiv:2607.08768v1 Announce Type: new Abstract: The rapid development of large language models and multimodal large language models has accelerated the emergence of proactive agents capable of operati

Unveiling Public Opinion: A Study of Sentiment Analysis Using LSTM and Traditional Models

ResearchDGX agent

arXiv:2607.07772v1 Announce Type: new Abstract: In this age of social media, sites like Twitter have become meeting places for people to share their views and feelings on a wide range of issues and cu

UtterTune: LoRA-Based Target-Language Pronunciation Edit and Control in Multilingual Text-to-Speech

ResearchDGX agent

arXiv:2508.09767v3 Announce Type: replace-cross Abstract: We propose UtterTune, a lightweight method for adapting a multilingual text-to-speech (TTS) system built on a large language model (LLM). It i

Validating LLMs in social science: Epistemic threats and emerging norms

SafetyDGX agent

arXiv:2607.07915v1 Announce Type: cross Abstract: Large language models (LLMs) are reshaping social science methodology. Researchers increasingly prompt language models to generate quantitative measur

When Debiasing Backfires: Counterintuitive Side Effects of Preprocessing-Based Stereotype Mitigation

ResearchDGX agent

arXiv:2607.07937v1 Announce Type: new Abstract: Preprocessing-based methods for stereotype mitigation, such as pre-/post-training on debiased corpora, are widely used in NLP. While these approaches re

XALPHA: A Memory-Driven AI Quant Researcher for Hypothesis-to-Code Alpha Discovery

SafetyDGX agent

arXiv:2607.08332v1 Announce Type: new Abstract: Financial markets are noisy, non-stationary, and high-dimensional, making it difficult to discover predictive and robust trading signals. Alpha discover

9 Jul 2026

A Word-Level Digital Reader of the Prasthanatrayi with Sankara's Bhasya: Corpus, Method, and an Open, Offline Reading Aid for the Advaita Vedanta Canon

ResearchDGX agent

arXiv:2607.07282v1 Announce Type: new Abstract: The Prasthanatrayi -- the ten principal Upanisads, the Brahmasutra, and the Bhagavadgita, with Sankara's commentaries (bhasya) -- is the foundational co

Behavior Leverage Imbalance in Multi-Teacher On-Policy Distillation

SafetyDGX agent

arXiv:2607.07050v1 Announce Type: new Abstract: Agentic language models must learn when to call tools, when to consume tool responses, and when to answer directly. This makes multi-teacher on-policy d

Billions of Sketches Reveal Hidden Cultural Variation in Human Concepts

ResearchDGX agent

arXiv:2607.07267v1 Announce Type: cross Abstract: Claims about the universality of human concepts have been predominantly assessed through linguistic similarity across languages and cultures. However,

C-DeltaTheta: Circuit-Restricted Weight Arithmetic for Selective Refusal

Local AiDGX agent

arXiv:2602.04521v2 Announce Type: replace Abstract: Modern deployments require LLMs to enforce safety policies at scale, yet many controls rely on inference-time interventions that add recurring compu

ContestTrade: A Multi-Agent Trading System Based on Internal Contest Mechanism

AgentsDGX agent

arXiv:2508.00554v4 Announce Type: replace-cross Abstract: In financial trading, large language model (LLM)-based agents demonstrate significant potential, but their decisions can be sensitive to noisy

DeLS-Spec: Decoupled Long-Short Contexts for Parallel Speculative Drafting

Local AiDGX agent

arXiv:2607.07409v1 Announce Type: new Abstract: Speculative decoding accelerates LLM inference by drafting multiple tokens and verifying them in parallel. Block-parallel drafters such as DFlash furthe

Dissociating the Internal Representations of Sycophancy in LLMs

ResearchDGX agent

arXiv:2607.07003v1 Announce Type: cross Abstract: Large Language Models (LLMs) frequently exhibit sycophancy, where they agree with a user's statement even when incorrect. While sycophancy is often tr

Does Bielik Know What It Doesn't Know? Activation Dispersion Separates Entity Familiarity from Factual Reliability Across Model Scale

ResearchDGX agent

arXiv:2607.07670v1 Announce Type: new Abstract: Large language models hallucinate most about entities they have never seen. We ask whether a model's activations betray entity familiarity before a sing

Dual Path Attribution: Efficient Attribution for SwiGLU-Transformers through Layer-Wise Target Propagation

ResearchDGX agent

arXiv:2603.19742v2 Announce Type: replace-cross Abstract: Understanding the internal mechanisms of transformer-based large language models (LLMs) is crucial for their reliable deployment and effective

Evaluating RAG Metrics in Applied Contexts: An Experiment, Its Findings and Its Limitations

ResearchDGX agent

arXiv:2607.07302v1 Announce Type: new Abstract: This paper reports an empirical study evaluating the relevance of several RAG metrics. The experiment is based on a question-answering dataset created b

Evaluation of Multilingual Ability to Use Spatial Deictic Expressions in Vision-Language Models

Model ReleasesDGX agent

arXiv:2607.07251v1 Announce Type: new Abstract: One of the expected abilities of vision-language models (VLMs) is spatial reasoning ability based on a given text and image. To evaluate the spatial rea

Fast, Slow, and Tool-augmented Thinking for LLMs: A Review

TutorialsDGX agent

arXiv:2508.12265v2 Announce Type: replace Abstract: Large Language Models (LLMs) have demonstrated remarkable progress in reasoning across diverse domains. However, effective reasoning in real-world t

Final Checkpoints Are Not Enough: Analyzing Latent Reasoning Faithfulness Along Training Trajectories

ResearchDGX agent

arXiv:2607.06648v1 Announce Type: cross Abstract: Latent reasoning methods perform multi-step inference entirely in the model's continuous hidden states, promising more compact and efficient reasoning

FourierQK: Spectral Preprocessing of Query-Key Projections Improves Transformer Attention

ResearchDGX agent

arXiv:2607.07478v1 Announce Type: cross Abstract: FFT-based spectral preprocessing of learned query-key (Q/K) projections substantially improves transformer attention on character-level language model

← Previous
1…2223242526…129
Next →