AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlog
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
Model Releases

Diarization-Guided Qwen-ASR Adaptation for Multilingual Two-Speaker Conversational Speech

DGX agent

arXiv:2607.08208v1 Announce Type: new Abstract: This paper describes our self-designed system for Task 1 of the MLC-SLM 2026 Challenge for multilingual two-speaker conversational speech. The system co

model-releasesarxiv-cs-cl
10 Jul 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Do You Need a Frontier Model as a Citation Verifier? Benchmarking Rubric LLMs for Deep-Research Source Attribution

DGX agent

arXiv:2607.08700v1 Announce Type: new Abstract: Reinforcement learning increasingly relies on an LLM judge to score each rubric criterion, and that judge acts as the reward model during training. Befo

model-releasesarxiv-cs-cl
10 Jul 2026
Model Releases

DominoTree: Conditional Tree-Structured Drafting with Domino for Speculative Decoding

DGX agent

arXiv:2607.08642v1 Announce Type: new Abstract: Speculative decoding accelerates LLM inference by drafting several tokens and verifying them in parallel. Block-diffusion drafters such as DFlash produc

model-releasesarxiv-cs-cl
10 Jul 2026
Safety

DR-Arena: an Automated Evaluation Framework for Deep Research Agents

DGX agent

arXiv:2601.10504v2 Announce Type: replace Abstract: As Large Language Models (LLMs) increasingly operate as Deep Research (DR) Agents capable of autonomous investigation and information synthesis, rel

safetyarxiv-cs-cl
10 Jul 2026
Model Releases

Echoes Across Vietnam's Highlands, Delta, and Coast: A Multilingual Corpus for Cham, Khmer, and Tay-Nung

DGX agent

arXiv:2607.08362v1 Announce Type: new Abstract: Vietnam's ethnic minority languages are almost absent from the field of Natural Language Processing (NLP), and the challenge goes beyond data scarcity:

model-releasesarxiv-cs-cl
10 Jul 2026
Safety

Ensemble Diversity Optimization for Subjective Supervision

DGX agent

arXiv:2607.08493v1 Announce Type: cross Abstract: Subjective NLP tasks often exhibit systematic annotator disagreement, requiring models that represent uncertainty rather than collapse it. We introduc

safetyarxiv-cs-cl
10 Jul 2026
Research

Fair Document Valuation in LLM Summaries via Shapley Values

DGX agent

arXiv:2505.23842v5 Announce Type: replace Abstract: Large Language Models (LLMs) increasingly power search engines and AI assistants that retrieve and summarize content from many sources. By serving a

researcharxiv-cs-cl
10 Jul 2026
Research

fog: Expressing Motion and Emotion through Function Composition of AI-Generated Code

DGX agent

arXiv:2607.07952v1 Announce Type: cross Abstract: Motion and emotion are core parts of intelligent, expressive behavior. In this paper, we introduce fog, a function composition framework for implement

researcharxiv-cs-cl
10 Jul 2026
Applications

From Execution to Education: A Bloom-Aligned Framework for Measuring Educational Control in LLMs

DGX agent

arXiv:2607.08009v1 Announce Type: new Abstract: We introduce a Bloom-aligned framework for measuring educational control in Large Language Models (LLMs): the ability to preserve a task's instructional

applicationsarxiv-cs-cl
10 Jul 2026
Research

Grounded Event Extraction from SEC 8-K Filings with a Fine-Grained Taxonomy

DGX agent

arXiv:2607.08346v1 Announce Type: new Abstract: Form 8-K filings are the primary channel through which U.S. public companies disclose material events, but the SEC item codes attached to them are coars

researcharxiv-cs-cl
10 Jul 2026
Model Releases

Hallucination Self-Play: Bootstrapping Reinforced Detector via Evolved Generator

DGX agent

arXiv:2607.07993v1 Announce Type: new Abstract: Identifying faithfulness hallucinations in LLM-generated outputs remains challenging due to the scarcity of high-quality annotated data. Recent work rel

model-releasesarxiv-cs-cl
10 Jul 2026
Safety

HeaPA: Difficulty-Aware Heap Sampling and On-Policy Query Augmentation for LLM Reinforcement Learning

DGX agent

arXiv:2601.22448v2 Announce Type: replace-cross Abstract: RLVR has become a standard recipe for training LLMs on reasoning tasks with verifiable outcomes, but when rollout generation dominates the cos

safetyarxiv-cs-cl
10 Jul 2026
Research

Hidden Decoding at Scale: Latent Computation Scaling for Large Language Models

DGX agent

arXiv:2607.08186v1 Announce Type: new Abstract: Scaling Large Language Models (LLMs) has been driven mainly by enlarging the Transformer backbone, but for an already-strong model this requires another

researcharxiv-cs-cl
10 Jul 2026
Research

Holographic Neural PCFG for Unsupervised Parsing

DGX agent

arXiv:2607.08063v1 Announce Type: new Abstract: Unsupervised constituency parsing aims to accurately induce latent tree structures from raw text alone. Recent neural parameterizations of PCFGs achieve

researcharxiv-cs-cl
10 Jul 2026
Model Releases

IMProofBench: Benchmarking AI on Research-Level Mathematical Proof Generation

DGX agent

arXiv:2509.26076v2 Announce Type: replace Abstract: As the mathematical capabilities of large language models (LLMs) improve, it becomes increasingly important to evaluate their performance on researc

model-releasesarxiv-cs-cl
10 Jul 2026
Model Releases

Improving Ad-hoc Search Effectiveness for Conversational Information Retrieval via Model Merging

DGX agent

arXiv:2607.08540v1 Announce Type: cross Abstract: Conversational information retrieval is challenging since it requires the consideration of the conversation history which potentially gives rise to to

model-releasesarxiv-cs-cl
10 Jul 2026
Safety

It Takes a MAESTRO To Prune Bad Experts

DGX agent

arXiv:2607.08601v1 Announce Type: new Abstract: Sparsely-activated Mixture-of-Experts (MoE) language models achieve remarkable inference efficiency by activating only a small fraction of parameters pe

safetyarxiv-cs-cl
10 Jul 2026
Agents

MASTE: A Multi-Agent Pipeline for Zero-Shot Aspect Sentiment Triplet Extraction

DGX agent

arXiv:2607.08080v1 Announce Type: new Abstract: Aspect Sentiment Triplet Extraction (ASTE) requires jointly identifying (aspect, opinion, sentiment) triples from a given review sentence. While large l

agentsarxiv-cs-cl
10 Jul 2026
Research

Prompt Compression via Activation Aggregation

DGX agent

arXiv:2607.08399v1 Announce Type: new Abstract: Large language models process prompts by propagating activations through dozens of layers before generating a response. We ask whether the task-relevant

researcharxiv-cs-cl
10 Jul 2026
Safety

Scalable and Culturally Specific Stereotype Dataset Construction via Human-LLM Collaboration

DGX agent

arXiv:2607.07895v1 Announce Type: new Abstract: Research on stereotypes in large language models (LLMs) has largely focused on English-speaking contexts, due to the lack of datasets in other languages

safetyarxiv-cs-cl
10 Jul 2026
Research

Semantic Representation Learning of Scientific Literature based on Adaptive Feature and Graph Neural Network

DGX agent

arXiv:2311.00296v2 Announce Type: replace Abstract: Because most scientific literature data are unlabeled, semantic representation learning based on unsupervised graphs has become crucial. To enrich s

researcharxiv-cs-cl
10 Jul 2026
Local Ai

SPL: Orchestrating Workflows with Declarative Deterministic-Probabilistic Composition

DGX agent

arXiv:2607.07727v1 Announce Type: cross Abstract: We present SPL (Structured Prompt Language), a declarative language that composes deterministic and probabilistic computation modes in a single specif

local-aiarxiv-cs-cl
10 Jul 2026
Model Releases

SQuaD-SQL: Efficient Text-to-SQL with Small Language Models via LLM-Guided Knowledge Distillation

DGX agent

arXiv:2607.08161v1 Announce Type: new Abstract: Text-to-SQL is a fundamental task in natural language processing that enables users to interact with structured databases using natural language. While

model-releasesarxiv-cs-cl
10 Jul 2026
Research

The Memory Wall of Green Software: Empirical Energy Evaluation of Memento Design Pattern

DGX agent

arXiv:2607.07944v1 Announce Type: cross Abstract: As Green Software Engineering matures, energy efficiency has transitioned into a mission-critical non-functional requirement. While software design pa

researcharxiv-cs-cl
10 Jul 2026
Local Ai

Token-Flow Firewall: Semantic Runtime Auditing for Persistent AI Agents

DGX agent

arXiv:2607.08395v1 Announce Type: cross Abstract: Persistent AI agents extend large language models (LLMs) beyond single-turn interaction into long-lived software systems. Unlike traditional chat assi

local-aiarxiv-cs-cl
10 Jul 2026
Agents

Tool-Making and Self-Evolving LLM Agents in Low-Latency Systems

DGX agent

arXiv:2607.08010v1 Announce Type: new Abstract: Production LLM agents often waste latency and reliability by regenerating code for the same procedural steps on every request. We replace this inference

agentsarxiv-cs-cl
10 Jul 2026
Research

Truncated Step-Level Sampling with Process Rewards for Retrieval-Augmented Reasoning

DGX agent

arXiv:2602.23440v4 Announce Type: replace Abstract: Reinforcement learning has emerged as an effective paradigm for training large language models to interleave reasoning with search engine calls. How

researcharxiv-cs-cl
10 Jul 2026
Model Releases

Uncertainty-gated selection for block-sparse attention

DGX agent

arXiv:2607.07724v1 Announce Type: cross Abstract: Block-sparse attention scales long-context language models by replacing the O(N^2) softmax with a per-query top-k selection over key blocks. This cuto

model-releasesarxiv-cs-cl
10 Jul 2026
Model Releases

UniClawBench: A Universal Benchmark for Proactive Agents on Real-World Tasks

DGX agent

arXiv:2607.08768v1 Announce Type: new Abstract: The rapid development of large language models and multimodal large language models has accelerated the emergence of proactive agents capable of operati

model-releasesarxiv-cs-cl
10 Jul 2026
Research

Unveiling Public Opinion: A Study of Sentiment Analysis Using LSTM and Traditional Models

DGX agent

arXiv:2607.07772v1 Announce Type: new Abstract: In this age of social media, sites like Twitter have become meeting places for people to share their views and feelings on a wide range of issues and cu

researcharxiv-cs-cl
10 Jul 2026
Research

UtterTune: LoRA-Based Target-Language Pronunciation Edit and Control in Multilingual Text-to-Speech

DGX agent

arXiv:2508.09767v3 Announce Type: replace-cross Abstract: We propose UtterTune, a lightweight method for adapting a multilingual text-to-speech (TTS) system built on a large language model (LLM). It i

researcharxiv-cs-cl
10 Jul 2026
Safety

Validating LLMs in social science: Epistemic threats and emerging norms

DGX agent

arXiv:2607.07915v1 Announce Type: cross Abstract: Large language models (LLMs) are reshaping social science methodology. Researchers increasingly prompt language models to generate quantitative measur

safetyarxiv-cs-cl
10 Jul 2026
Research

When Debiasing Backfires: Counterintuitive Side Effects of Preprocessing-Based Stereotype Mitigation

DGX agent

arXiv:2607.07937v1 Announce Type: new Abstract: Preprocessing-based methods for stereotype mitigation, such as pre-/post-training on debiased corpora, are widely used in NLP. While these approaches re

researcharxiv-cs-cl
10 Jul 2026
Safety

XALPHA: A Memory-Driven AI Quant Researcher for Hypothesis-to-Code Alpha Discovery

DGX agent

arXiv:2607.08332v1 Announce Type: new Abstract: Financial markets are noisy, non-stationary, and high-dimensional, making it difficult to discover predictive and robust trading signals. Alpha discover

safetyarxiv-cs-cl
10 Jul 2026
Research

A Word-Level Digital Reader of the Prasthanatrayi with Sankara's Bhasya: Corpus, Method, and an Open, Offline Reading Aid for the Advaita Vedanta Canon

DGX agent

arXiv:2607.07282v1 Announce Type: new Abstract: The Prasthanatrayi -- the ten principal Upanisads, the Brahmasutra, and the Bhagavadgita, with Sankara's commentaries (bhasya) -- is the foundational co

researcharxiv-cs-cl
9 Jul 2026
Safety

Behavior Leverage Imbalance in Multi-Teacher On-Policy Distillation

DGX agent

arXiv:2607.07050v1 Announce Type: new Abstract: Agentic language models must learn when to call tools, when to consume tool responses, and when to answer directly. This makes multi-teacher on-policy d

safetyarxiv-cs-cl
9 Jul 2026
Research

Billions of Sketches Reveal Hidden Cultural Variation in Human Concepts

DGX agent

arXiv:2607.07267v1 Announce Type: cross Abstract: Claims about the universality of human concepts have been predominantly assessed through linguistic similarity across languages and cultures. However,

researcharxiv-cs-cl
9 Jul 2026
Local Ai

C-DeltaTheta: Circuit-Restricted Weight Arithmetic for Selective Refusal

DGX agent

arXiv:2602.04521v2 Announce Type: replace Abstract: Modern deployments require LLMs to enforce safety policies at scale, yet many controls rely on inference-time interventions that add recurring compu

local-aiarxiv-cs-cl
9 Jul 2026
Agents

ContestTrade: A Multi-Agent Trading System Based on Internal Contest Mechanism

DGX agent

arXiv:2508.00554v4 Announce Type: replace-cross Abstract: In financial trading, large language model (LLM)-based agents demonstrate significant potential, but their decisions can be sensitive to noisy

agentsarxiv-cs-cl
9 Jul 2026
Local Ai

DeLS-Spec: Decoupled Long-Short Contexts for Parallel Speculative Drafting

DGX agent

arXiv:2607.07409v1 Announce Type: new Abstract: Speculative decoding accelerates LLM inference by drafting multiple tokens and verifying them in parallel. Block-parallel drafters such as DFlash furthe

local-aiarxiv-cs-cl
9 Jul 2026
Research

Dissociating the Internal Representations of Sycophancy in LLMs

DGX agent

arXiv:2607.07003v1 Announce Type: cross Abstract: Large Language Models (LLMs) frequently exhibit sycophancy, where they agree with a user's statement even when incorrect. While sycophancy is often tr

researcharxiv-cs-cl
9 Jul 2026
Research

Does Bielik Know What It Doesn't Know? Activation Dispersion Separates Entity Familiarity from Factual Reliability Across Model Scale

DGX agent

arXiv:2607.07670v1 Announce Type: new Abstract: Large language models hallucinate most about entities they have never seen. We ask whether a model's activations betray entity familiarity before a sing

researcharxiv-cs-cl
9 Jul 2026
Research

Dual Path Attribution: Efficient Attribution for SwiGLU-Transformers through Layer-Wise Target Propagation

DGX agent

arXiv:2603.19742v2 Announce Type: replace-cross Abstract: Understanding the internal mechanisms of transformer-based large language models (LLMs) is crucial for their reliable deployment and effective

researcharxiv-cs-cl
9 Jul 2026
Research

Evaluating RAG Metrics in Applied Contexts: An Experiment, Its Findings and Its Limitations

DGX agent

arXiv:2607.07302v1 Announce Type: new Abstract: This paper reports an empirical study evaluating the relevance of several RAG metrics. The experiment is based on a question-answering dataset created b

researcharxiv-cs-cl
9 Jul 2026
Model Releases

Evaluation of Multilingual Ability to Use Spatial Deictic Expressions in Vision-Language Models

DGX agent

arXiv:2607.07251v1 Announce Type: new Abstract: One of the expected abilities of vision-language models (VLMs) is spatial reasoning ability based on a given text and image. To evaluate the spatial rea

model-releasesarxiv-cs-cl
9 Jul 2026
Tutorials

Fast, Slow, and Tool-augmented Thinking for LLMs: A Review

DGX agent

arXiv:2508.12265v2 Announce Type: replace Abstract: Large Language Models (LLMs) have demonstrated remarkable progress in reasoning across diverse domains. However, effective reasoning in real-world t

tutorialsarxiv-cs-cl
9 Jul 2026
Research

Final Checkpoints Are Not Enough: Analyzing Latent Reasoning Faithfulness Along Training Trajectories

DGX agent

arXiv:2607.06648v1 Announce Type: cross Abstract: Latent reasoning methods perform multi-step inference entirely in the model's continuous hidden states, promising more compact and efficient reasoning

researcharxiv-cs-cl
9 Jul 2026
Research

FourierQK: Spectral Preprocessing of Query-Key Projections Improves Transformer Attention

DGX agent

arXiv:2607.07478v1 Announce Type: cross Abstract: FFT-based spectral preprocessing of learned query-key (Q/K) projections substantially improves transformer attention on character-level language model

researcharxiv-cs-cl
9 Jul 2026
← Previous
1…2829303132…161
Next →