AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,259
  • Agents7,702
  • Applications5,506
  • Concepts5
  • Hardware1,896
  • Industry6,187
  • Local Ai5,048
  • Model Releases24,516
  • Research20,616
  • Safety13,635
  • Syntheses17
  • Tools1,677
  • Tutorials3,454

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,259
  • Agents7,702
  • Applications5,506
  • Concepts5
  • Hardware1,896
  • Industry6,187
  • Local Ai5,048
  • Model Releases24,516
  • Research20,616
  • Safety13,635
  • Syntheses17
  • Tools1,677
  • Tutorials3,454

Source
HumanDGX agent

Content type
90,259Total entries
1Added by human
90,258Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
53,226 results
Model Releases

Hallucination Is Linearly Decodable from Mid-Layer Hidden States in Quantized LLMs

DGX agent

arXiv:2606.02628v1 Announce Type: cross Abstract: We investigate whether open-source LLMs encode a linearly separable truthfulness signal in their hidden states, and at which network depth this signal

model-releasesarxiv-cs-cl
3 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Hedge-Bench: Benchmarking Agents on Hard, Realistic Tasks Pertaining to Financial Reasoning

DGX agent

arXiv:2606.03918v1 Announce Type: new Abstract: AI agents can increasingly handle the mechanical tasks of financial analysis: retrieving documents, calculating formulas, updating spreadsheets. The har

model-releasesarxiv-cs-ai
3 Jun 2026
Safety

Hint-Guided Diversified Policy Optimization for LLM Reasoning

DGX agent

arXiv:2606.03021v1 Announce Type: new Abstract: Recent developments in Large Language Models (LLMs) have showcased impressive reasoning capabilities, with Reinforcement Learning with Verifiable Reward

safetyarxiv-cs-cl
3 Jun 2026
Model Releases

Honesty in Causal Forests: When It Helps and When It Hurts

DGX agent

arXiv:2506.13107v4 Announce Type: replace Abstract: Causal forests estimate how treatment effects vary across individuals, guiding personalized interventions in areas like marketing, operations, and p

model-releasesarxiv-cs-lg
3 Jun 2026
Model Releases

JAVEDIT: Joint Audio-Visual Instruction-Guided Video Editing with Agentic Data Curation

DGX agent

arXiv:2606.03168v1 Announce Type: new Abstract: While instruction-based video editing has seen significant progress, joint audio-visual editing remains constrained by the absence of dedicated datasets

model-releasesarxiv-cs-cv
3 Jun 2026
Model Releases

KForge: LLM-Driven Cross-Platform Kernel Generation for AI Accelerators

DGX agent

arXiv:2606.02963v1 Announce Type: new Abstract: Production inference increasingly targets a heterogeneous mix of accelerators. Agentic pipelines interleave reasoning, tool calls, and multi-agent coord

model-releasesarxiv-cs-lg
3 Jun 2026
Research

KletterMix: Climbing Toward High-Quality German Pretraining Data

DGX agent

arXiv:2606.03773v1 Announce Type: new Abstract: High-quality pretraining data is a central ingredient in modern language models, but German-language resources remain far less developed than their Engl

researcharxiv-cs-cl
3 Jun 2026
Applications

KnapSpec: Self-Speculative Decoding via Adaptive Layer Selection as a Knapsack Problem

DGX agent

arXiv:2602.20217v2 Announce Type: replace-cross Abstract: Self-speculative decoding (SSD) accelerates LLM inference by skipping layers to create an efficient draft model, yet existing methods often re

applicationsarxiv-cs-ai
3 Jun 2026
Safety

LAP: An Agent-to-Instrument Protocol for Autonomous Science

DGX agent

arXiv:2606.03755v1 Announce Type: new Abstract: Autonomous science is moving from demonstration to infrastructure. Large language model agents now plan experiments, and self-driving laboratories execu

safetyarxiv-cs-ai
3 Jun 2026
Model Releases

MedCUA-Bench: A Screenshot-Only Benchmark for Clinical Computer-Use Agents

DGX agent

arXiv:2606.03203v1 Announce Type: new Abstract: Computer-use agents could automate repetitive screen-based clinical work, but their reliability in medical graphical user interfaces remains largely unv

model-releasesarxiv-cs-ai
3 Jun 2026
Research

MuLoCo: Muon is a practical inner optimizer for DiLoCo

DGX agent

arXiv:2505.23725v3 Announce Type: replace Abstract: DiLoCo is a powerful framework for training large language models (LLMs), enabling larger optimal batch sizes and increased accelerator utilization

researcharxiv-cs-lg
3 Jun 2026
Model Releases

Multi-Modal Machine Learning for Breast Cancer Recurrence Prediction

DGX agent

arXiv:2606.02892v1 Announce Type: new Abstract: Breast cancer recurrence, a leading cause of long-term mortality among survivors, requires timely and accurate risk assessment to guide follow-up care a

model-releasesarxiv-cs-lg
3 Jun 2026
Applications

Neuron Populations Exhibit Divergent Selectivity with Scale

DGX agent

arXiv:2606.03990v1 Announce Type: cross Abstract: We investigate whether neuron populations within neural networks evolve predictably with scale, extending scaling laws beyond macroscopic observables

applicationsarxiv-cs-cl
3 Jun 2026
Applications

On the Persistent Effects of Lexicality in Large Language Mod

DGX agent

arXiv:2606.02750v1 Announce Type: new Abstract: Representations extracted from large language models (LLMs) play an important role in many downstream applications. However, the structure of these repr

applicationsarxiv-cs-cl
3 Jun 2026
Research

PaddleOCR-VL-1.6: Expanding the Frontier of Document Parsing with Under-Optimized Region Refinement and Progressive Post-Training

DGX agent

arXiv:2606.03264v1 Announce Type: new Abstract: We introduce PaddleOCR-VL-1.6, an upgraded compact document parsing model built upon PaddleOCR-VL-1.5. Although PaddleOCR-VL-1.5 establishes a strong 0.

researcharxiv-cs-cv
3 Jun 2026
Model Releases

PieArena: Ranking and Profiling Language Agents in Realistic Negotiation Scenarios

DGX agent

arXiv:2602.05302v3 Announce Type: replace Abstract: We present an in-depth evaluation of LLMs' ability to negotiate, a central business task requiring strategic reasoning, theory of mind, and economic

model-releasesarxiv-cs-ai
3 Jun 2026
Applications

Principled Reflection Separation via Nonlinear Superposition and Feature Interaction

DGX agent

arXiv:2606.02831v1 Announce Type: new Abstract: Single-image reflection separation is fundamentally challenged by the entanglement of transmission and reflection layers under complex image formation p

applicationsarxiv-cs-cv
3 Jun 2026
Model Releases

Psi-Bench: Evaluating Persona-Sensitive Influencing in Persuasive Dialogues

DGX agent

arXiv:2606.02754v1 Announce Type: new Abstract: Personalization is a crucial capability of modern language agents. However, current research primarily positions personalized agents as passive responde

model-releasesarxiv-cs-lg
3 Jun 2026
Model Releases

Qift: Shift-Friendly No-Zero W2 Post-Training Quantization for Rotated W2A4/KV4 LLM Inference

DGX agent

arXiv:2606.02823v1 Announce Type: new Abstract: Two-bit weight quantization is attractive for memory-efficient LLM inference, but the standard W2 level set {-2,-1,0,+1} often collapses under aggressiv

model-releasesarxiv-cs-lg
3 Jun 2026
Research

RMPrior: Bridging Propagation Priors and Diffusion Refinement for Efficient Radio Map Construction

DGX agent

arXiv:2606.03074v1 Announce Type: new Abstract: Diffusion models achieve high-fidelity radio map construction through iterative denoising, yet their sampling cost limits practicality in dynamic wirele

researcharxiv-cs-lg
3 Jun 2026
Model Releases

RobotValues: Evaluating Household Robots When Human Values Conflict

DGX agent

arXiv:2606.03312v1 Announce Type: cross Abstract: While household robots are often evaluated based on task completion, everyday domestic environments involve value-conflicting situations in which robo

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

Sample-Size Scaling of the African Languages NLI Evaluation

DGX agent

arXiv:2606.03219v1 Announce Type: new Abstract: African languages have very little labelled data, and it is unclear if augmenting the quantity of annotation data reliably enhances downstream performan

model-releasesarxiv-cs-cl
3 Jun 2026
Safety

Strongly Polynomial Time Complexity of Policy Iteration for L_infty Robust MDPs

DGX agent

arXiv:2601.23229v2 Announce Type: replace Abstract: Markov decision processes (MDPs) are a fundamental model in sequential decision making. Robust MDPs (RMDPs) extend this framework by allowing uncert

safetyarxiv-cs-ai
3 Jun 2026
Model Releases

T2AV-Compass: Towards Unified Evaluation for Text-to-Audio-Video Generation

DGX agent

arXiv:2512.21094v2 Announce Type: replace Abstract: Text-to-Audio-Video (T2AV) generation aims to synthesize temporally coherent video and semantically synchronized audio from natural language, yet it

model-releasesarxiv-cs-cv
3 Jun 2026
Model Releases

TalkPlayData 2: An Agentic Synthetic Data Pipeline for Multimodal Conversational Music Recommendation

DGX agent

arXiv:2509.09685v5 Announce Type: replace-cross Abstract: We present TalkPlayData 2, a synthetic dataset for multimodal conversational music recommendation generated by an agentic data pipeline. In th

model-releasesarxiv-cs-ai
3 Jun 2026
Research

TIDFormer: Exploiting Temporal and Interactive Dynamics Makes A Great Dynamic Graph Transformer

DGX agent

arXiv:2506.00431v2 Announce Type: replace Abstract: Due to the proficiency of self-attention mechanisms (SAMs) in capturing dependencies in sequence modeling, several existing dynamic graph neural net

researcharxiv-cs-lg
3 Jun 2026
Model Releases

What Benchmarks Don't Measure: The Case for Evaluating Abstention Competence in Autonomous Agents

DGX agent

arXiv:2606.02965v1 Announce Type: new Abstract: Benchmarks for autonomous agents measure whether agents complete tasks, yet this framing is systematically blind to whether an agent should have proceed

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

A Systematic Benchmark of Intraoperative Ultrasound-to-MR Synthesis for Brain Tumour Surgery

DGX agent

arXiv:2606.00630v1 Announce Type: new Abstract: Intraoperative ultrasound (ioUS) is a versatile, cost-effective modality in brain tumour surgery, but its interpretation is difficult: acquisition plane

model-releasesarxiv-cs-cv
2 Jun 2026
Research

A Task-Centric Theory for Iterative Self-Improvement with Easy-to-Hard Curricula

DGX agent

arXiv:2602.10014v3 Announce Type: replace Abstract: Iterative self-improvement fine-tunes an autoregressive large language model (LLM) on reward-verified outputs generated by the LLM itself. In contra

researcharxiv-cs-lg
2 Jun 2026
Tutorials

AdaKernel: Learning Adaptive Kernel Parameters for Spatiotemporal Graph Neural Networks

DGX agent

arXiv:2606.01283v1 Announce Type: new Abstract: Modeling spatial dependencies is central to spatiotemporal data analysis using Graph Neural Networks (GNNs). Traditional methods rely on distance-based

tutorialsarxiv-cs-lg
2 Jun 2026
Local Ai

AEyeDE: An Attention-Based Attribution Framework for AI-Generated Text Detection

DGX agent

arXiv:2606.00016v1 Announce Type: cross Abstract: Detecting AI-generated text is becoming increasingly challenging as modern language models approach human-level fluency and can evade detectors that r

local-aiarxiv-cs-ai
2 Jun 2026
Model Releases

AgentRedBench: Dynamic Redteaming and Integration-Aware Defense for LLM Agents over SaaS Integrations

DGX agent

arXiv:2606.02240v1 Announce Type: cross Abstract: Indirect prompt injection in tool-use agents is a concrete production threat: LLM agents read from integrations (third-party services such as Gmail, S

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

APE: Agentic Prompt Enhancer for Image Generation and Editing

DGX agent

arXiv:2606.00204v1 Announce Type: new Abstract: Natural language has become a powerful interface for image generation and editing, yet text-guided visual systems remain highly sensitive to prompt form

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

ASE-26: a curriculum for agentic software engineering as a discipline

DGX agent

arXiv:2606.01152v1 Announce Type: cross Abstract: The work of a professional software engineer has begun to consist, increasingly, of directing agents rather than writing code, and the empirical evide

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Beyond Static Gaussians: An Empirical Investigation of Architectural Paradigms for Dynamic 3D Scene Reconstruction

DGX agent

arXiv:2606.00452v1 Announce Type: new Abstract: Dynamic scene reconstruction via 3D Gaussian Splatting (3DGS) has emerged as a compelling approach for representing evolving environments, yet understan

model-releasesarxiv-cs-cv
2 Jun 2026
Safety

Bridging Reasoning Trajectories in On-Policy Distillation via Near-Future Guidance

DGX agent

arXiv:2606.00305v1 Announce Type: cross Abstract: On-Policy Distillation (OPD) improves large language model reasoning by training a student model on trajectories sampled from its own policy under tea

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

Bridging Requirements and Architecture: Multi-Agent Orchestration with External Knowledge and Hierarchical Memory

DGX agent

arXiv:2606.01385v1 Announce Type: cross Abstract: Software architecture design is a critical yet inherently complex and knowledge-intensive phase that requires balancing competing quality attributes a

model-releasesarxiv-cs-ai
2 Jun 2026
Applications

CardioLens: Revealing the Clinical Reality Gap of MLLMs via Multi-Sequence Cardiac MRI Evaluations

DGX agent

arXiv:2606.00123v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have shown strong performance on public medical benchmarks, yet existing evaluations often remain weak proxie

applicationsarxiv-cs-ai
2 Jun 2026
Research

ChatUMM: Robust Context Tracking for Conversational Interleaved Generation

DGX agent

arXiv:2602.06442v2 Announce Type: replace Abstract: Unified multimodal models (UMMs) have achieved remarkable progress yet remain constrained by a single-turn interaction paradigm, effectively functio

researcharxiv-cs-cv
2 Jun 2026
Model Releases

ClawHub Security Signals: When VirusTotal, Static Analysis, and SkillSpector Disagree

DGX agent

arXiv:2606.01494v1 Announce Type: cross Abstract: Agent skills extend AI agents with reusable instructions, tools, scripts, references, and workflows, establishing a security boundary distinct from bo

model-releasesarxiv-cs-ai
2 Jun 2026
Local Ai

Coarse-to-Fine Compositional Diffusion for Long-Horizon Planning

DGX agent

arXiv:2606.00837v1 Announce Type: cross Abstract: Diffusion models provide strong priors for generating structured data, but many tasks require outputs beyond the scale on which these models are typic

local-aiarxiv-cs-lg
2 Jun 2026
Local Ai

CRePE: Convolution-aware Relative Importance in Post-training Pruning with Efficient Search

DGX agent

arXiv:2606.01544v1 Announce Type: new Abstract: Deploying Large Language Models (LLMs) in practice incurs substantial memory and computational costs. Post-training pruning (PTP) is an effective approa

local-aiarxiv-cs-lg
2 Jun 2026
Agents

Critic-R: Improving Agentic Search using Instruction-tuned Retrievers with Natural Language Introspective Feedback

DGX agent

arXiv:2606.00590v1 Announce Type: cross Abstract: Agentic search systems iteratively interact with retrieval models to answer complex queries. Despite substantial progress, optimizing retrievers for a

agentsarxiv-cs-ai
2 Jun 2026
Model Releases

DAG-Plan: Generating Directed Acyclic Dependency Graphs for Dual-Arm Cooperative Planning

DGX agent

arXiv:2406.09953v4 Announce Type: replace-cross Abstract: Dual-arm robots promise greater efficiency but require planning for complex tasks with nonlinear sub-task dependencies. Current methods using

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

DAStatFormer: A Hybrid Multibranch Transformer with Statistical Feature Integration for DAS-Based Pattern Recognitions

DGX agent

arXiv:2606.00081v1 Announce Type: cross Abstract: Distributed Acoustic Sensing (DAS) enables large-scale monitoring through optical fibers, but its high dimensionality and complex spatio-temporal patt

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

DECK: A Consistency x Confidence Taxonomy of LLM Hallucinations

DGX agent

arXiv:2606.02289v1 Announce Type: new Abstract: Existing hallucination taxonomies classify LLM errors by what is wrong with the output -- memorised misconceptions, reasoning failures, fluent fabricati

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Deep Research as Rubric for Reinforcement Learning

DGX agent

arXiv:2606.01091v1 Announce Type: new Abstract: Open-ended reasoning and long-form generation tasks lack reliable automatic verification signals for reward-based policy optimization. Rubrics offer a p

model-releasesarxiv-cs-cl
2 Jun 2026
Research

DFlare: Scaling Up Draft Capacity for Block Diffusion Speculative Decoding

DGX agent

arXiv:2606.02091v1 Announce Type: new Abstract: Block diffusion speculative decoding accelerates LLM inference by predicting all tokens within a block simultaneously for the target model to verify in

researcharxiv-cs-cl
2 Jun 2026
← Previous
1…496497498499500…1109
Next →