AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlog
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,646 results
Safety

Deep Research Pretraining via Predictive Navigation

DGX agent

arXiv:2608.00432v1 Announce Type: new Abstract: Deep research agents are often trained on expensive, environment-grounded tool-use trajectories that require repeated retrieval, document inspection, an

safetyarxiv-cs-cl
4 Aug 2026
Model Releases

DeepSurvey-Bench: Evaluating Academic Value of Automatically Generated Scientific Surveys

X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

arXiv:2601.15307v2 Announce Type: replace-cross Abstract: The rapid development of automated survey generation technology has made it increasingly important to establish a comprehensive benchmark to e

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

DeltaFlow: Noise-Adaptive Bidirectional Gated Delta Networks for Embedded Language Flows

DGX agent

arXiv:2608.01240v1 Announce Type: new Abstract: Embedded Language Flows (ELF) rely primarily on full non-causal attention for iterative denoising, repeatedly incurring quadratic sequence-mixing cost a

model-releasesarxiv-cs-cl
4 Aug 2026
Tutorials

Dense Language Generation Made Simple: Deterministic, Randomized, and Multi-Order Algorithms

DGX agent

arXiv:2608.01320v1 Announce Type: cross Abstract: Language generation in the limit is a theoretical framework for studying how a generator can learn to produce new valid strings from a stream of posit

tutorialsarxiv-cs-cl
4 Aug 2026
Agents

Diagnosing Search Behavior and Failure Modes in Long-Horizon Search Agents

DGX agent

arXiv:2608.01913v1 Announce Type: cross Abstract: Deep search agents answer difficult information-seeking questions by iteratively issuing search queries to gather supporting evidence, but it remains

agentsarxiv-cs-cl
4 Aug 2026
Model Releases

DiffusionGemma Technical Report

DGX agent

arXiv:2608.00146v1 Announce Type: new Abstract: We introduce DiffusionGemma, an experimental open-weight language model that uses discrete diffusion to generate text at exceptionally high speed. Rathe

model-releasesarxiv-cs-cl
4 Aug 2026
Research

Discriminative Axis, Not Data Volume: What a Contrastive Corpus Teaches an Audio Embedding

DGX agent

arXiv:2608.01560v1 Announce Type: new Abstract: Scaling the corpus is the default remedy when a contrastive representation lacks an attribute. We report a case where it does nothing, and identify what

researcharxiv-cs-cl
4 Aug 2026
Safety

Disentangled Contrastive Learning for Zero-Shot Multilingual Dense Retrieval

DGX agent

arXiv:2608.02189v1 Announce Type: cross Abstract: Multilingual dense retrieval aims to handle queries and documents across different languages based on a unified retriever model. The challenge lies in

safetyarxiv-cs-cl
4 Aug 2026
Safety

Distill Where You Fail: Recovering Learning Signals of Negative RL-Groups from Adaptive Teacher Guidance

DGX agent

arXiv:2608.00782v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has become a standard paradigm for post-training large language models (LLMs). While Group Relativ

safetyarxiv-cs-cl
4 Aug 2026
Model Releases

Divergent large language model predictions from convergent representations in ambiguous word pairs

DGX agent

arXiv:2608.01816v1 Announce Type: new Abstract: In this work we investigate how decoder-only transformers resolve lexical ambiguity through layer-by-layer analysis of three models spanning three param

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

DLLM-TTS: Block Discrete Diffusion Language Model for Text-to-Speech Synthesis

DGX agent

arXiv:2608.00011v1 Announce Type: new Abstract: Current text-to-speech systems face a trade-off: autoregres- sive codec language models produce highly intelligible speech but require large-scale model

model-releasesarxiv-cs-cl
4 Aug 2026
Agents

DocNavRAG: Document-Structured Graph RAG with Stateful Evidence Construction for Complex Document Question Answering

DGX agent

arXiv:2608.01565v1 Announce Type: new Abstract: Answering complex questions over large document collections requires assembling complementary evidence across sections and documents. GraphRAG offers st

agentsarxiv-cs-cl
4 Aug 2026
Research

Does Accuracy Equal Evidence? Reasoning Faithfulness under KV Cache Compression

DGX agent

arXiv:2608.01631v1 Announce Type: new Abstract: KV cache compression is commonly evaluated by final-answer accuracy, implicitly assuming that preserving the answer also preserves the reasoning that su

researcharxiv-cs-cl
4 Aug 2026
Tutorials

Does Machine 'know' interpersonal pragmatics? Evidence from MARBERT's learning of emoji pragmatics in Arabic digital discourse

DGX agent

arXiv:2608.01174v1 Announce Type: new Abstract: This study examines Transformer-based models' ability to learn emoji pragmatics in Arabic digital discourse (ADD), providing evidence from MARBERT's beh

tutorialsarxiv-cs-cl
4 Aug 2026
Applications

Does the Competitive Component of Adversarial Self-Play Improve Legal Reasoning? A Controlled Negative Result

DGX agent

arXiv:2608.01559v1 Announce Type: cross Abstract: Adversarial self-play is an appealing recipe for legal reasoning: have a student model draft an argument, have an adversary attack it, and reward the

applicationsarxiv-cs-cl
4 Aug 2026
Model Releases

Domain-Specific Evaluation of Text-to-Speech Systems: A Multi-Metric Benchmarking Study

DGX agent

arXiv:2608.02235v1 Announce Type: new Abstract: Recent advances in neural text-to-speech (TTS) systems have substantially improved speech naturalness and intelligibility across many languages. However

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

Don't Judge a Book by its Cover: Testing LLMs' Robustness Under Logical Obfuscation

DGX agent

arXiv:2602.01132v2 Announce Type: replace Abstract: Tasks such as solving arithmetic equations, evaluating truth tables, and completing syllogisms are handled well by large language models (LLMs) in t

model-releasesarxiv-cs-cl
4 Aug 2026
Applications

Don't Offer What Can't Be Done: Deterministic Executability Gating for LLM Skill Selection at Scale

DGX agent

arXiv:2608.01050v1 Announce Type: cross Abstract: Production LLM agents that select from large skill libraries face a limitation that semantic relevance alone cannot resolve: a skill may match a user'

applicationsarxiv-cs-cl
4 Aug 2026
Applications

Douyin Multimodal Embedding Model Technical Report

DGX agent

arXiv:2608.02148v1 Announce Type: cross Abstract: Multimodal representation learning is a cornerstone of modern AI. By encoding multimodal queries and targets into vectors, it powers industrial search

applicationsarxiv-cs-cl
4 Aug 2026
Research

EHR2Path: Comprehensive Pathway-Level Modeling of Longitudinal Patient Trajectories from Multimodal Electronic Health Records

DGX agent

arXiv:2506.04831v3 Announce Type: replace-cross Abstract: Forecasting how a patient's condition is likely to evolve, including possible deterioration, recovery, treatment needs, and care transitions,

researcharxiv-cs-cl
4 Aug 2026
Research

Enhancing Large Language Model Reasoning with Reward Models: An Analytical Survey

DGX agent

arXiv:2510.01925v3 Announce Type: replace Abstract: Reward models (RMs) play a critical role in enhancing the reasoning performance of LLMs. For example, they can provide training signals to finetune

researcharxiv-cs-cl
4 Aug 2026
Tutorials

Entity-Faithful Repair of Synthetic Supervision for Zero-Shot Image Captioning

DGX agent

arXiv:2608.00994v1 Announce Type: cross Abstract: Zero-shot image captioning aims to generate image descriptions without annotated image-text pairs. Recent approaches exploit text-to-image models to s

tutorialsarxiv-cs-cl
4 Aug 2026
Model Releases

ET-Prune: Evidence-Aware Dynamic Budgeting for Visual Token Pruning in Text-Rich MLLMs

DGX agent

arXiv:2608.01979v1 Announce Type: cross Abstract: Visual token pruning reduces the inference cost of multimodal large language models, but a fixed token ratio is poorly matched to text-rich inputs. In

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

Evaluating VLMs on Multimodal Aristotelian Persuasion Tasks

DGX agent

arXiv:2608.01238v1 Announce Type: new Abstract: Vision Language Models (VLMs) have demonstrated exceptional performance across various tasks. However, they have not yet been thoroughly evaluated on mo

model-releasesarxiv-cs-cl
4 Aug 2026
Local Ai

EviSD: Evidence-Conditioned Self-Distillation for Search-Augmented Agents

DGX agent

arXiv:2608.01359v1 Announce Type: new Abstract: Outcome-based reinforcement learning enables search-augmented language agents to learn from verifiable final answers, but its trajectory-level credit ca

local-aiarxiv-cs-cl
4 Aug 2026
Tutorials

Exemplars in Disguise: Pure Exemplar Models Mimic Abstraction-First Learning

DGX agent

arXiv:2608.00821v1 Announce Type: new Abstract: Whether idiosyncratic, item-specific knowledge is learned before abstract class-level generalizations, or vice versa, is a central question in language

tutorialsarxiv-cs-cl
4 Aug 2026
Safety

Expert-Choice Routing Enables Adaptive Computation in Diffusion Language Models

DGX agent

arXiv:2604.01622v2 Announce Type: replace-cross Abstract: Diffusion language models (DLMs) enable parallel, non-autoregressive text generation, yet existing DLM mixture-of-experts (MoE) models inherit

safetyarxiv-cs-cl
4 Aug 2026
Safety

Exploiting Intrinsic Duality for Multi-Hop Question Generation

DGX agent

arXiv:2608.00712v1 Announce Type: new Abstract: Multi hop question generation (MQG) aims to generate questions from multiple given documents and target answers, whereas question answering (QA) focuses

safetyarxiv-cs-cl
4 Aug 2026
Research

Exploring More to Solve More: Boosting Diversity in Text Diffusion Models via Entropy-Based Guidance

DGX agent

arXiv:2608.00024v1 Announce Type: new Abstract: Although diffusion models have revolutionized continuous domains like image synthesis through high quality generations and controllable guidance mechani

researcharxiv-cs-cl
4 Aug 2026
Model Releases

Fast and Accurate Quotation Attribution in Literary Texts

DGX agent

arXiv:2608.02359v1 Announce Type: new Abstract: Attributing quotations to their speakers in literary texts remains an open challenge. Standard methods, which independently predict a speaker mention fo

model-releasesarxiv-cs-cl
4 Aug 2026
Research

Fenced Citation-Context Retrieval for Case Law: Temporal Leakage and Degree Control Across Two Jurisdictions

DGX agent

arXiv:2607.17142v3 Announce Type: replace-cross Abstract: Prior case retrieval (PCR) aims to identify the precedent cases relevant to the facts of a query case. Incoming citation context, the text wit

researcharxiv-cs-cl
4 Aug 2026
Research

Few-Shot Biomedical Relation Extraction with Large Language Models: A Viable Alternative to Supervised Learning?

DGX agent

arXiv:2606.15412v2 Announce Type: replace Abstract: Biomedical relation extraction (BioRE) is a key step in transforming biomedical literature into structured knowledge. Most existing approaches rely

researcharxiv-cs-cl
4 Aug 2026
Model Releases

FinHardBench: Can LLMs Generate Latency-Aware Hardware for Financial Computing?

DGX agent

arXiv:2608.00909v1 Announce Type: new Abstract: Can large language models generate not just correct, but fast hardware? This paper investigates the question in financial FPGA design, where 5-10 nanose

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

Floor, Ceiling, and the Fusion Gap: How Much of Crowd Reading Attention Can Machines Predict?

DGX agent

arXiv:2608.01704v1 Announce Type: cross Abstract: A benchmark score means nothing without knowing what a trivial method achieves and what the best possible method could achieve. We construct both boun

model-releasesarxiv-cs-cl
4 Aug 2026
Research

From Chains to Trees: Parent-Conditioned Drafting for Semi-Autoregressive Speculative Decoding

DGX agent

arXiv:2608.02123v1 Announce Type: new Abstract: Speculative decoding accelerates LLM inference only when drafted continuations survive target-model verification. Semi-autoregressive drafters such as D

researcharxiv-cs-cl
4 Aug 2026
Model Releases

From Direction to Magnitude: How Multimodal Instruction-Tuning Reorganizes the Geometric Encoding of Identity-Specifying Prompts in Transformer Hidden States

DGX agent

arXiv:2607.09842v2 Announce Type: replace-cross Abstract: We investigate whether identity-specifying system prompts produce statistically distinguishable geometric fingerprints in the hidden-state tra

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

From We to Me: Theory Informed Narrative Shift with Abductive Reasoning

DGX agent

arXiv:2603.03320v2 Announce Type: replace Abstract: Effective communication often relies on aligning a message with an audience's narrative and worldview. Narrative shift involves transforming text to

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

Gaokerena: A Small Persian Medical Language Model Family

DGX agent

arXiv:2608.00932v1 Announce Type: new Abstract: The integration of artificial intelligence into medical question-answering systems has advanced rapidly; however, research remains predominantly focused

model-releasesarxiv-cs-cl
4 Aug 2026
Research

Geometry-Guided Layerwise FFN Width Allocation in Transformers

DGX agent

arXiv:2608.02064v1 Announce Type: cross Abstract: Feed-forward networks (FFNs) account for a large fraction of Transformer parameters, yet their hidden width is usually constant across depth. We ask w

researcharxiv-cs-cl
4 Aug 2026
Model Releases

Global Optimization and Inference-Time Region Grafting for Agentic Workflows

DGX agent

arXiv:2608.02353v1 Announce Type: new Abstract: Recent advances in agentic workflow optimization automate workflow design through task-specific workflow search or input-conditioned architecture select

model-releasesarxiv-cs-cl
4 Aug 2026
Research

GradCuit: Credit-Assigned Gradient Flow Enables Robust and Interpretable Test-Time Latent Reasoning

DGX agent

arXiv:2608.02585v1 Announce Type: cross Abstract: Optimization-based latent reasoning improves large language model outputs by optimizing instance-specific continuous states at test time while keeping

researcharxiv-cs-cl
4 Aug 2026
Research

HAFI-VLM: A Frequency Perspective for Diagnosing and Enhancing Visual Perception in Vision-Language Models

DGX agent

arXiv:2608.02124v1 Announce Type: cross Abstract: Vision-language models (VLMs) remain unreliable when predictions require fine-grained visual evidence. We identify a previously overlooked cause: spec

researcharxiv-cs-cl
4 Aug 2026
Model Releases

HarnessCompass: Guiding Automatic Harness Evolution toward Generalizable and Effective Agent Harnesses

DGX agent

arXiv:2608.01918v1 Announce Type: cross Abstract: Harness design plays a critical role in agent performance by shaping how large language models (LLMs) perceive, reason over, and act within executable

model-releasesarxiv-cs-cl
4 Aug 2026
Hardware

HetRoute Heterogeneous and Cost-aware Collaborative Routing Framework for Distributed Edge MoE Inference

DGX agent

arXiv:2608.00577v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) models have become a dominant architecture for large-scale AI services, yet deploying them over geo-distributed heterogeneous

hardwarearxiv-cs-cl
4 Aug 2026
Safety

Hierarchical Pre-Training of Vision Encoders with Large Language Model

DGX agent

arXiv:2604.00086v2 Announce Type: replace-cross Abstract: The field of computer vision has experienced significant advancements through scalable vision encoders and multimodal pre-training frameworks.

safetyarxiv-cs-cl
4 Aug 2026
Model Releases

HopRefusalBench: Diagnosing Refusal Failures in Search-Augmented Agents for Multi-Hop Reasoning

DGX agent

arXiv:2608.01358v1 Announce Type: new Abstract: Search-augmented large language model agents are increasingly capable of solving knowledge-intensive tasks, but their behavior when a multi-hop question

model-releasesarxiv-cs-cl
4 Aug 2026
Safety

Human-LLM Alignment in Language Attitudes Toward Non-Native Japanese

DGX agent

arXiv:2608.01629v1 Announce Type: new Abstract: Large language models (LLMs) increasingly evaluate human writing in high-stakes domains such as hiring and academic assessment, putting non-native speak

safetyarxiv-cs-cl
4 Aug 2026
Applications

Hylog: A Hybrid Approach to Logging Text Production in Non-alphabetic Scripts

DGX agent

arXiv:2601.17753v2 Announce Type: replace Abstract: Research keyloggers are essential for cognitive studies of text production, yet most fail to capture the on-screen transformations performed by Inpu

applicationsarxiv-cs-cl
4 Aug 2026
← Previous
1…1011121314…160
Next →