AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
91,060Total entries
1Added by human
91,059Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,793 results
Local Ai

Spatial-Omni: Spatial Audio Understanding Integration in Multimodal LLMs via FOA Encoding

DGX agent

arXiv:2606.10738v1 Announce Type: cross Abstract: Recent multimodal large language models mainly process audio as monaural signals, thereby discarding the spatial cues contained in spatial audio for s

local-aiarxiv-cs-ai
10 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

STORM: Stepwise Token Optimization with Reward-Guided Beam Search

DGX agent

arXiv:2606.10621v1 Announce Type: cross Abstract: Modern retrieval increasingly relies on dense and learned-sparse neural models that are effective but require encoding the entire corpus into a specia

researcharxiv-cs-ai
10 Jun 2026
Safety

Task Robustness via Re-Labelling Vision-Action Robot Data

DGX agent

arXiv:2606.10918v1 Announce Type: cross Abstract: The recent trend in scaling models for robot learning has resulted in impressive policies that can perform various manipulation tasks and generalize t

safetyarxiv-cs-lg
10 Jun 2026
Safety

TD-Grokking: Learning from Zero-Reward Problems by Training-Time Decomposition

DGX agent

arXiv:2606.09883v1 Announce Type: cross Abstract: Large language models (LLMs) have made remarkable progress in reasoning tasks, largely driven by post-training paradigms, especially reinforcement lea

safetyarxiv-cs-ai
10 Jun 2026
Agents

The Arbiter Agent: Continually Monitoring Multi-Agent Conversations to Detect Emergent Misalignment

DGX agent

arXiv:2606.10747v1 Announce Type: new Abstract: As AI systems built from multiple language-model agents become more common, they are increasingly used to make decisions together: discussing, negotiati

agentsarxiv-cs-ai
10 Jun 2026
Tutorials

There is a lot of justified anger at Anthropic for sandbagging Fable 5 for AI development tasks. But an unanticipated side effect is that th…

DGX agent

There is a lot of justified anger at Anthropic for sandbagging Fable 5 for AI development tasks. But an unanticipated side effect is that third-party evaluators can no longer credibly use the model fo

tutorialsjeremy-howard--x
10 Jun 2026
Agents

Toward Secure LLM Agents: Threat Surfaces, Attacks, Defenses, and Evaluation

DGX agent

arXiv:2606.10749v1 Announce Type: cross Abstract: Large language model (LLM) agents are rapidly moving from conversational interfaces to software components that plan, invoke tools, maintain memory, a

agentsarxiv-cs-ai
10 Jun 2026
Model Releases

UMI-Bench 1.0: An Open and Reproducible Real-World Benchmark for Tabletop Robotic Manipulation with UMI Data

DGX agent

arXiv:2606.10382v1 Announce Type: new Abstract: Real-robot evaluation is essential for understanding whether learned manipulation policies can operate reliably outside curated demonstrations. This nee

model-releasesarxiv-cs-ro
10 Jun 2026
Research

Uncertainty-aware Multi-fidelity Closure via Conditional Normalizing Flows

DGX agent

arXiv:2606.09857v1 Announce Type: new Abstract: Reduced-order models (ROMs) provide an efficient surrogate for complex multiscale systems, but their predictive accuracy is often compromised by truncat

researcharxiv-cs-lg
10 Jun 2026
Safety

UniPET: a universal network for high-quality PET image denoising across varied dose reduction factors

DGX agent

arXiv:2606.11131v1 Announce Type: new Abstract: Most existing deep learning-based PET image denoising methods assume a fixed and known dose reduction factor (DRF) for low-dose PET images. However, the

safetyarxiv-cs-cv
10 Jun 2026
Model Releases

What makes a harness a harness: necessary and sufficient conditions for an agent harness

DGX agent

arXiv:2606.10106v1 Announce Type: cross Abstract: The term agent harness now circulates widely in software engineering with generative artificial intelligence. It names the layer that wraps a language

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

What Matters in Orchestrating Robot Policies: A Systematic Study of Hierarchical VLA Agents

DGX agent

arXiv:2606.10267v1 Announce Type: cross Abstract: Hierarchical vision-language-action (Hi-VLA) systems have emerged as a promising paradigm for complex robot manipulation, by using high-level VLM plan

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

When Design Rules Break: Benchmark Composition Determines Whether Label Informativeness Predicts GNN Aggregator Choice

DGX agent

arXiv:2606.10249v1 Announce Type: new Abstract: We examine whether graph neural network (GNN) design rules generalize across benchmark families by studying aggregator selection (sum, mean, max) on 24

model-releasesarxiv-cs-lg
10 Jun 2026
Model Releases

WHU-Infra3D: A Full-stack Multi-modal Dataset and Benchmark for 3D Roadside Infrastructure Inventory

DGX agent

arXiv:2606.09882v1 Announce Type: new Abstract: The paradigm of digital twin cities is shifting from coarse visual mapping toward more precise and actionable digitization of urban assets. However, exi

model-releasesarxiv-cs-cv
10 Jun 2026
Model Releases

With SoftBank apparently struggling to get a margin loan against it’s OpenAI shares, it’s maybe time to repost this, from two years ago. The…

DGX agent

With SoftBank apparently struggling to get a margin loan against it’s OpenAI shares, it’s maybe time to repost this, from two years ago. The vast majority of my earlier worries remain: 9 reasons that

model-releasesgary-marcus--x
10 Jun 2026
Tutorials

A Systematic Study of Behavioral Cloning for Scientific Data Annotation

DGX agent

arXiv:2606.07568v1 Announce Type: cross Abstract: Scientific data annotation, such as tracking animals in video or proofreading neural reconstructions, remains bottlenecked by the 'last mile' problem:

tutorialsarxiv-cs-ai
9 Jun 2026
Tutorials

Adversarial Attack and Disturbance Detection by Hadamard-Coded Output Representations for Object Detection and Semantic Segmentation

DGX agent

arXiv:2606.09536v1 Announce Type: new Abstract: Conventional one-hot encodings often yield poorly calibrated models, being overconfident under attack, and letting entropy-based detection algorithms fa

tutorialsarxiv-cs-cv
9 Jun 2026
Local Ai

AGENTSERVESIM: A Hardware-aware Simulator for Multi-Turn LLM Agent Serving

DGX agent

arXiv:2606.09613v1 Announce Type: cross Abstract: Multi-turn LLM agents interleave model calls with external tool invocations, shifting serving from stateless request processing to stateful program ex

local-aiarxiv-cs-ai
9 Jun 2026
Model Releases

AgroOmni: A Large-Scale Multi-view Agricultural Dataset for Cross-Scale Multimodal Reasoning

DGX agent

arXiv:2603.14342v2 Announce Type: replace-cross Abstract: Modern agricultural data is sourced from diverse platforms and spans multiple spatial scales, ranging from ground-level close-up photography t

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Anthropic says Claude Fable 5 uses conservative safety classifiers that trigger a fallback to Claude Opus 4.8 in <5% of sessions, in areas like cybersecurity (Anthropic)

DGX agent

Anthropic: Anthropic says Claude Fable 5 uses conservative safety classifiers that trigger a fallback to Claude Opus 4.8 in <5% of sessions, in areas like cybersecurity — Today we're launching Claude

model-releasestechmeme
9 Jun 2026
Model Releases

Anthropic says Fable 5 is available on Pro, Max, Team, and seat-based Enterprise plans through June 22, after which using Fable 5 will require usage credits (Rebecca Bellan/TechCrunch)

DGX agent

Rebecca Bellan / TechCrunch: Anthropic says Fable 5 is available on Pro, Max, Team, and seat-based Enterprise plans through June 22, after which using Fable 5 will require usage credits — Anthropic is

model-releasestechmeme
9 Jun 2026
Tutorials

ATLAS: Verifier-Guided Adaptive Latent Activation Steering for Efficient LLM Reasoning

DGX agent

arXiv:2601.03093v2 Announce Type: replace Abstract: Recent work on activation and latent steering has demonstrated that modifying internal representations can effectively guide large language models (

tutorialsarxiv-cs-lg
9 Jun 2026
Local Ai

Attention Illuminates LLM Reasoning: The Preplan-and-Anchor Rhythm Enables Fine-Grained Policy Optimization

DGX agent

arXiv:2510.13554v2 Announce Type: replace-cross Abstract: The reasoning pattern of Large language models (LLMs) remains opaque, and reinforcement learning (RL) typically applies uniform credit across

local-aiarxiv-cs-lg
9 Jun 2026
Research

Audio-FLAN: An Instruction-Following Dataset for Unified Audio Understanding and Generation of Speech, Music, and Sound

DGX agent

arXiv:2502.16584v2 Announce Type: replace-cross Abstract: Recent advancements in audio tokenization have significantly enhanced the integration of audio capabilities into large language models (LLMs).

researcharxiv-cs-ai
9 Jun 2026
Model Releases

Auditable Graph-Guided Root Cause Analysis for Kubernetes Incidents

DGX agent

arXiv:2606.08590v1 Announce Type: cross Abstract: Kubernetes incidents are diagnosed reliably only when a root-cause system's reported gains come from incident evidence rather than scenario-specific s

model-releasesarxiv-cs-ai
9 Jun 2026
Safety

Baichuan-M4: A Clinical-Grade Medical Agent System for Continuous Care

DGX agent

arXiv:2606.08982v1 Announce Type: new Abstract: Baichuan-M4 is Baichuan Intelligence's clinical-grade medical large model, designed for continuous care rather than single-turn medical question answeri

safetyarxiv-cs-ai
9 Jun 2026
Model Releases

Bayesian Selective Latent Inference for Wastewater-First Influenza Monitoring

DGX agent

arXiv:2606.09433v1 Announce Type: new Abstract: Wastewater influenza surveillance can reveal community circulation before clinical reporting, but wastewater alone is not a fully identifiable proxy for

model-releasesarxiv-cs-ai
9 Jun 2026
Safety

Beyond Scalar Rewards by Internalizing Reasoning into Score Distributions

DGX agent

arXiv:2606.09076v1 Announce Type: new Abstract: Reward models are central to text-to-image post-training, but visual preference is subjective and better represented as a distribution over rubric score

safetyarxiv-cs-cv
9 Jun 2026
Model Releases

BioVid: Autoregressive Video Generation with Biological Behavior Semantic Comprehension

DGX agent

arXiv:2606.08674v1 Announce Type: cross Abstract: Existing video generation frameworks treat sequence duration as an externally prescribed parameter -- fixed frame counts or text prompts -- producing

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Breaking the Bubble: Asynchronous Pipeline Parallel Training with Bounded Weight Inconsistency

DGX agent

arXiv:2606.07881v1 Announce Type: new Abstract: Pipeline parallelism is essential for training large neural networks, but existing schedules trade off throughput, memory, and optimization consistency.

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

Bridged SBI: Correcting Biased Low-Fidelity Posteriors for Cost-Efficient High-Fidelity Inference

DGX agent

arXiv:2606.09155v1 Announce Type: new Abstract: Accurate calibration of particle-based simulators is crucial for robotic earthwork simulation, but analytical calibration is challenging due to this tas

model-releasesarxiv-cs-ro
9 Jun 2026
Model Releases

C3VD-DEFCOL: A Deformable Colonoscopy Dataset with Time-Resolved 3D Ground Truth and Realistic Appearance

DGX agent

arXiv:2606.07891v1 Announce Type: new Abstract: 3D reconstruction could improve colonoscopy by estimating mucosal coverage and alerting clinicians to missed regions during screening. However, algorith

model-releasesarxiv-cs-cv
9 Jun 2026
Research

CAMF-Det: Closure-Aware Multimodal Fusion for LiDAR-Camera 3D Object Detection on UAV Platforms

DGX agent

arXiv:2606.09143v1 Announce Type: new Abstract: Multimodal 3D object detection based on LiDAR and cameras has demonstrated excellent performance in ground-vehicle scenarios, but has not been explored

researcharxiv-cs-cv
9 Jun 2026
Model Releases

Can You Trust What You See? Human and AI Detection of Synthetic Legal Evidence

DGX agent

arXiv:2606.07613v1 Announce Type: cross Abstract: Visual evidence has long been treated as a reliable form of legal proof, but advances in artificial intelligence (AI) are undermining that assumption.

model-releasesarxiv-cs-ai
9 Jun 2026
Research

CausShield: Sample Reconstruction-Resilient Vertical FL via Causal Representation Learning

DGX agent

arXiv:2606.08027v1 Announce Type: cross Abstract: Vertical federated learning (VFL) is a distributed learning paradigm that leverages vertically partitioned features across isolated parties without sh

researcharxiv-cs-ai
9 Jun 2026
Model Releases

Claude Fable 5 hands-on: impressive results working on complex projects, like building an interactive isochrone map or a data analysis tool in just 9.5 hours (Ethan Mollick/One Useful Thing)

DGX agent

Ethan Mollick / One Useful Thing: Claude Fable 5 hands-on: impressive results working on complex projects, like building an interactive isochrone map or a data analysis tool in just 9.5 hours — Claude

model-releasestechmeme
9 Jun 2026
Model Releases

@claudeai please buy more data centers asap

DGX agent

Jerry Liu's post urges Anthropic to rapidly expand its data center infrastructure to support Claude AI's growing demand and operational needs. The post reflects concerns about scaling computational re

model-releasesjerry-liu--x
9 Jun 2026
Model Releases

Closing the Sim-to-Real Gap: An Evaluation Framework for Autonomous Cyber Defense Configuration of Commercial EDR

DGX agent

arXiv:2606.08168v1 Announce Type: cross Abstract: Leading commercial endpoint detection and response (EDR) products have shifted from operator-configured rule sets to multi-component systems where aut

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Condition-Gated Reasoning for Context-Dependent Biomedical Question Answering

DGX agent

arXiv:2602.17911v3 Announce Type: replace-cross Abstract: Current biomedical question answering (QA) systems often assume that medical knowledge applies uniformly, yet real-world clinical reasoning is

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Conditional Normalizing Flows for Forward and Backward Joint State and Parameter Estimation

DGX agent

arXiv:2601.07013v2 Announce Type: replace-cross Abstract: Traditional filtering algorithms for state estimation -- such as classical Kalman filtering, unscented Kalman filtering, and particle filters

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

ContextShift: A Controlled Benchmark for Context Dependence in Object Detection

DGX agent

arXiv:2606.09495v1 Announce Type: new Abstract: Modern object detectors achieve strong performance on standard benchmarks, yet their robustness to contextual variation remains insufficiently understoo

model-releasesarxiv-cs-cv
9 Jun 2026
Local Ai

Continuous Language Diffusion as a Decoder-Interface Problem

DGX agent

arXiv:2606.08810v1 Announce Type: cross Abstract: Gaussian-corrupted sentence embeddings have no direct linguistic interpretation, yet continuous diffusion language models can generate fluent text fro

local-aiarxiv-cs-lg
9 Jun 2026
Model Releases

Convergence Bound and Critical Batch Size of Muon Optimizer

DGX agent

arXiv:2507.01598v5 Announce Type: replace Abstract: Muon, a recently proposed optimizer that leverages the inherent matrix structure of neural network parameters, has demonstrated strong empirical per

model-releasesarxiv-cs-lg
9 Jun 2026
Safety

Correct Looks Better: Pairwise Comparisons Reveal Accuracy Rankings

DGX agent

arXiv:2606.09409v1 Announce Type: new Abstract: Pairwise comparisons combined with aggregation methods like Elo have become central to evaluating generative models, yet concerns remain that they rewar

safetyarxiv-cs-ai
9 Jun 2026
Model Releases

DAL-PCQA: Enabling Distortion-Level and Language-Driven Reasoning for Point Cloud Quality Assessment

DGX agent

arXiv:2606.07938v1 Announce Type: new Abstract: Point Cloud Quality Assessment (PCQA) methods typically predict scalar Mean Opinion Scores (MOS), which quantify overall perceptual degradation but do n

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

Decision-Aware Memory Cards: Counterfactual-Inspired Context Selection and Compression for Tool-Using LLM Agents

DGX agent

arXiv:2606.08151v1 Announce Type: new Abstract: Tool-using LLM agents often fail not because relevant text is absent, but because decisive evidence is not selected, compressed, or surfaced at action t

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

DHAuDS: A Dynamic and Heterogeneous Audio Benchmark for Test-Time Adaptation

DGX agent

arXiv:2511.18421v2 Announce Type: replace-cross Abstract: Existing Test-time Adaptation (TTA) studies rely heavily on static and homogeneous corruption protocols, such as ImageNet-C and CIFAR-10-C/100

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

Distortion-Aware PETR for BEV Object Detection with Mixed Pinhole-Fisheye Cameras

DGX agent

arXiv:2606.08680v1 Announce Type: new Abstract: Fisheye cameras are widely deployed in autonomous driving perception suites for their low cost and full-coverage field of view (FOV), yet their potentia

model-releasesarxiv-cs-cv
9 Jun 2026
← Previous
1…695696697698699…1371
Next →