AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,338
  • Agents7,718
  • Applications5,508
  • Concepts5
  • Hardware1,906
  • Industry6,195
  • Local Ai5,052
  • Model Releases24,549
  • Research20,616
  • Safety13,637
  • Syntheses17
  • Tools1,678
  • Tutorials3,457

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,338
  • Agents7,718
  • Applications5,508
  • Concepts5
  • Hardware1,906
  • Industry6,195
  • Local Ai5,052
  • Model Releases24,549
  • Research20,616
  • Safety13,637
  • Syntheses17
  • Tools1,678
  • Tutorials3,457

Source
HumanDGX agent

Content type
AllBlog
90,338Total entries
1Added by human
90,337Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,187 results
Safety

GPO: Learning from Critical Steps to Improve LLM Reasoning

DGX agent

arXiv:2509.16456v3 Announce Type: replace Abstract: Large language models (LLMs) are increasingly used in various domains, showing impressive potential on different tasks. Recently, reasoning LLMs hav

safetyarxiv-cs-ai
11 Jun 2026
Model Releases

GraphInfer-Bench: Benchmarking LLM's Inference Capability on Graphs

X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

arXiv:2606.11562v1 Announce Type: cross Abstract: Graph analysis underlies many applications whose answers cannot be looked up in a single record or retrieved along a path: laundering rings, drug repu

model-releasesarxiv-cs-cl
11 Jun 2026
Model Releases

Holding the FP8 Quality Ceiling at 8-Bit Weights and Activations: INT8 and GGUF Post-Training Quantization of Ideogram 4.0 for Consumer GPUs

DGX agent

arXiv:2606.12280v1 Announce Type: new Abstract: Post-training quantization lets large text-to-image diffusion transformers run on consumer GPUs, yet the hardware-specific trade-offs are seldom measure

model-releasesarxiv-cs-lg
11 Jun 2026
Safety

IAPO: Input Attribution-Aware Policy Optimization for Tool Use in Small Multimodal Agents

DGX agent

arXiv:2606.11652v1 Announce Type: new Abstract: This paper investigates reinforcement learning (RL) methods for improving tool-calling capabilities in multimodal small language model (SLM) agents. Whi

safetyarxiv-cs-lg
11 Jun 2026
Safety

Improving Generalization and Data Efficiency with Diffusion in Offline Multi-agent RL

DGX agent

arXiv:2307.01472v2 Announce Type: replace Abstract: We present a novel Diffusion Offline Multi-agent Model (DOM2) for offline Multi-Agent Reinforcement Learning (MARL). Different from existing algorit

safetyarxiv-cs-ai
11 Jun 2026
Research

Judging Against the Reference: Uncovering Knowledge-Driven Failures in LLM-Judges on QA Evaluation

DGX agent

arXiv:2601.07506v2 Announce Type: replace Abstract: While large language models (LLMs) are increasingly used as automatic judges for question answering (QA) and other reference-conditioned evaluation

researcharxiv-cs-cl
11 Jun 2026
Tutorials

Learning Patterns and Abstractions from Perceptual Sequences

DGX agent

arXiv:2503.10973v2 Announce Type: replace Abstract: Cognition swiftly breaks high-dimensional sensory streams into familiar parts and uncovers their relations. Why do structures emerge, and how do the

tutorialsarxiv-cs-lg
11 Jun 2026
Model Releases

LibriConvo: Simulating Conversations from Read Literature for ASR and Diarization

DGX agent

arXiv:2510.23320v2 Announce Type: replace-cross Abstract: We introduce LibriConvo, a synthetic conversational speech corpus for speaker diarization and automatic speech recognition (ASR), built by ins

model-releasesarxiv-cs-cl
11 Jun 2026
Agents

Measuring Epistemic Resilience of LLMs Under Misleading Medical Context

DGX agent

arXiv:2606.12291v1 Announce Type: new Abstract: Large language models (LLMs) now reach expert-level scores on medical licensing exams, encouraging the assumption that high scores imply safe medical ju

agentsarxiv-cs-cl
11 Jun 2026
Model Releases

MedCTA: A Benchmark for Clinical Tool Agents

DGX agent

arXiv:2606.11702v1 Announce Type: cross Abstract: To make clinically grounded decisions, medical AI agents are expected to go beyond simple recognition and be capable of tool retrieval, evidence acqui

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

MemNovo: Look Back at the Spectrum for Balanced De Novo Peptide Sequencing from Mass Spectrometry

DGX agent

arXiv:2606.11868v1 Announce Type: new Abstract: De novo peptide sequencing from tandem mass spectrometry is pivotal in proteomics, enabling identification of novel peptides without reference databases

model-releasesarxiv-cs-lg
11 Jun 2026
Model Releases

Metadata-Aware Multi-Prompt Reasoning for Zero-Shot Accident Understanding

DGX agent

arXiv:2606.12047v1 Announce Type: cross Abstract: In this paper, we address the problem of zero-shot understanding of accidents from surveillance videos by identifying when an impact event occurs, wha

model-releasesarxiv-cs-ai
11 Jun 2026
Local Ai

MODF-SIR: A Multi-agent Omni-modal Distilled Framework for Social Intelligence Reasoning

DGX agent

arXiv:2606.12018v1 Announce Type: new Abstract: We propose a multi-agent collaborative framework built upon a lightweight Multimodal Large Language Model (MLLM), specifically designed for social intel

local-aiarxiv-cs-ai
11 Jun 2026
Hardware

MPK: A Compiler and Runtime for Mega-Kernelizing Tensor Programs

DGX agent

arXiv:2512.22219v2 Announce Type: replace-cross Abstract: We introduce Mirage Persistent Kernel (MPK), the first compiler and runtime system that automatically transforms multi-GPU model inference int

hardwarearxiv-cs-lg
11 Jun 2026
Model Releases

Multi-Agent Reasoning with Adaptive Worker Allocation for Stance Detection

DGX agent

arXiv:2606.11609v1 Announce Type: new Abstract: Stance detection requires identifying an author's position toward a target, often from short-form texts where stance is implicit, indirect, or rhetorica

model-releasesarxiv-cs-cl
11 Jun 2026
Model Releases

NetBurst: Event-Centric Forecasting of Bursty, Intermittent Time Series

DGX agent

arXiv:2510.22397v2 Announce Type: replace-cross Abstract: Network operators monitor their infrastructure by collecting telemetry data such as packet counts, byte rates, or flow volumes, yet answering

model-releasesarxiv-cs-lg
11 Jun 2026
Agents

OmniBioTwin: A System-of-Twinned-Systems Framework for Health Digital Twins

DGX agent

arXiv:2606.11264v1 Announce Type: cross Abstract: Health digital twins (HDTs) promise patient-specific modeling and decision support but current approaches remain structurally fragmented: monolithic m

agentsarxiv-cs-ai
11 Jun 2026
Model Releases

Phase Transitions in Attention: A Bayesian Theory of Copy Head Emergence

DGX agent

arXiv:2606.12058v1 Announce Type: cross Abstract: Attention is the key mechanism underlying in-context learning in transformers, and attention patterns have been observed empirically to emerge abruptl

model-releasesarxiv-cs-lg
11 Jun 2026
Safety

Plan-and-Verify Video Reward Reasoning with Spatio-Temporal Scene Graph Grounding

DGX agent

arXiv:2606.11838v1 Announce Type: new Abstract: Reward models for text-to-video (T2V) generation guide post-training but often fail at fine-grained semantic alignment. We trace this to two structural

safetyarxiv-cs-cv
11 Jun 2026
Model Releases

Pop quiz, which of these is no longer true (or at least directionally true), two years later?

DGX agent

Pop quiz, which of these is no longer true (or at least directionally true), two years later? 9 reasons that OpenAI could someday be seen as the WeWork of AI: 👉 Lots of competitors are catching up. 👉

model-releasesgary-marcus--x
11 Jun 2026
Model Releases

Precision-Aware Illumination-Disentangled Vision Transformer for Spacecraft 6D Pose Estimation

DGX agent

arXiv:2606.11619v1 Announce Type: new Abstract: Vision sensors provide a lightweight solution for spacecraft proximity operations, but monocular spacecraft 6D pose estimation remains difficult under i

model-releasesarxiv-cs-cv
11 Jun 2026
Safety

Redesign Mixture-of-Experts Routers with Manifold Power Iteration

DGX agent

arXiv:2606.12397v1 Announce Type: cross Abstract: Router is the cornerstone component to the Mixture-of-Experts models. Serving as expert proxies, the rows of the router matrix compute their similarit

safetyarxiv-cs-ai
11 Jun 2026
Research

Renewable Lasso without Batch-Number Constraints: A Gradient-Enhanced Approach

DGX agent

arXiv:2606.11738v1 Announce Type: cross Abstract: We study online estimation for high-dimensional generalized linear models with streaming data. First, for the non-distributed setting, we propose a gr

researcharxiv-cs-lg
11 Jun 2026
Model Releases

Restless bandits with imperfect binary feedback: PCL-indexability analysis and computation

DGX agent

arXiv:2606.11192v1 Announce Type: new Abstract: We study restless bandits with binary latent states and imperfect binary feedback, motivated by opportunistic spectrum access with sensing errors. For t

model-releasesarxiv-cs-lg
11 Jun 2026
Model Releases

RoVE: Rotary Value Embeddings Attention for Relative Position-dependent Value Pathways

DGX agent

arXiv:2606.11275v1 Announce Type: cross Abstract: Rotary Position Embeddings (RoPE) make attention scores position-relative but leave the value pathway position-blind: the message sent by a value toke

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

SceneMiner: Identity-Preserving Multi-Task Fine-Tuning for Unified BEV Scene Mining

DGX agent

arXiv:2606.11507v1 Announce Type: new Abstract: Mining hard, safety-critical scenes from driving logs is bottlenecked by the absence of difficulty labels, and no single proxy, collision risk, trajecto

model-releasesarxiv-cs-cv
11 Jun 2026
Model Releases

SheafStain: Sheaf-Theoretic Schrodinger Bridge for Spatially and Biologically Coherent Virtual Staining

DGX agent

arXiv:2606.11846v1 Announce Type: new Abstract: Current virtual staining approaches offer the potential for time- and cost-efficient biomarker quantification in cancer diagnostics and prognostics. How

model-releasesarxiv-cs-cv
11 Jun 2026
Safety

Signed Compression Progress on a Sealed Audit is Goodhart-Resistant

DGX agent

arXiv:2606.11417v1 Announce Type: cross Abstract: Compression progress is a long-standing proposal for intrinsic motivation: reward an agent when its world model becomes better at predicting or compre

safetyarxiv-cs-ai
11 Jun 2026
Model Releases

Spatially Coupled Phase-to-Depth Calibration for Fringe Projection Profilometry

DGX agent

arXiv:2606.11601v1 Announce Type: new Abstract: In fringe projection profilometry (FPP), depth is commonly recovered by fitting a phase-to-depth relation independently at each camera pixel. Although s

model-releasesarxiv-cs-cv
11 Jun 2026
Research

SpikeDecoder: Realizing the GPT Architecture with Spiking Neural Networks

DGX agent

arXiv:2606.12287v1 Announce Type: cross Abstract: The Transformer architecture is widely regarded as the most powerful tool for natural language processing, but due to a high number of complex operati

researcharxiv-cs-ai
11 Jun 2026
Research

Temporal2Seq: A Unified Framework for Temporal Video Understanding Tasks

DGX agent

arXiv:2409.18478v2 Announce Type: replace Abstract: With the development of video understanding, there is a proliferation of tasks for clip-level temporal video analysis, including temporal action det

researcharxiv-cs-cv
11 Jun 2026
Safety

The Art of Interrogation: Consistency Amplifies Factuality in Spatial Reasoning

DGX agent

arXiv:2606.11918v1 Announce Type: new Abstract: Current Large Reasoning Models (LRMs) exhibit remarkable general capabilities but significantly underperform in spatial reasoning tasks. Existing approa

safetyarxiv-cs-ai
11 Jun 2026
Model Releases

Toward Generalist Autonomous Research via Hypothesis-Tree Refinement

DGX agent

arXiv:2606.11926v1 Announce Type: cross Abstract: Scientific progress depends on a repeated loop of exploration, experimentation, and abstraction. Researchers test candidate directions, interpret the

model-releasesarxiv-cs-ai
11 Jun 2026
Research

Vector Quantized Latent Concepts: A Scalable Alternative to Clustering-Based Concept Discovery

DGX agent

arXiv:2602.02726v2 Announce Type: replace-cross Abstract: Large language models (LLMs) encode rich semantic information in their hidden states, yet it remains difficult to understand what information

researcharxiv-cs-cl
11 Jun 2026
Model Releases

Verifiable Environments Are LEGO Bricks: Recursive Composition for Reasoning Generalization

DGX agent

arXiv:2606.12373v1 Announce Type: new Abstract: Reinforcement Learning (RL) with verifiable environments has emerged as a powerful approach for enhancing the reasoning capabilities of Large Language M

model-releasesarxiv-cs-cl
11 Jun 2026
Model Releases

VICX: Generalizable Robot Manipulation via Video Generation and In-Context Operator Network

DGX agent

arXiv:2606.12028v1 Announce Type: new Abstract: Generalizable robot manipulation requires not only task-level reasoning over unseen scenes, but also reliable grounding of visual plans into embodiment-

model-releasesarxiv-cs-ro
11 Jun 2026
Research

Vision Transformers for Face Recognition Need More Registers

DGX agent

arXiv:2606.12036v1 Announce Type: new Abstract: Recent advances in Vision Transformers (ViTs) for face recognition (FR) have moved beyond the standard CLS-token paradigm. In this paradigm, a special c

researcharxiv-cs-cv
11 Jun 2026
Research

A Controlled Audit of Pretraining Contamination in Public Medical Vision-Language Benchmarks

DGX agent

arXiv:2606.10066v1 Announce Type: cross Abstract: Medical vision-language models (VLMs) are evaluated on public benchmarks whose images and question-answer pairs have been freely downloadable for year

researcharxiv-cs-ai
10 Jun 2026
Safety

A Reliable Fault Diagnosis Method Based on Belief Rule Base Consider Robustness Analysis

DGX agent

arXiv:2606.10500v1 Announce Type: new Abstract: In equipment operation, the implementation of fault diagnosis is essential to ensure the continuity and safety of production equipment, improve operatio

safetyarxiv-cs-ai
10 Jun 2026
Applications

Active-Passive Federated Learning for Vertically Partitioned Multi-view Data

DGX agent

arXiv:2409.04111v2 Announce Type: replace Abstract: Vertical federated learning is a natural and elegant approach to integrate multi-view data vertically partitioned across devices (clients) while pre

applicationsarxiv-cs-lg
10 Jun 2026
Model Releases

ASA: Backbone-Training-Free Representation Engineering for Tool-Calling Agents

DGX agent

arXiv:2602.04935v3 Announce Type: replace-cross Abstract: Adapting LLM agents to domain-specific tool calling remains notably brittle under evolving interfaces. Prompt and schema engineering is easy t

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Attention Amnesia in Hybrid LLMs: When CoT Fine-Tuning Breaks Long-Range Recall, and How to Fix It

DGX agent

arXiv:2606.11052v1 Announce Type: new Abstract: Chain-of-thought (CoT) supervised fine-tuning (SFT) is widely adopted to improve reasoning ability, yet we find that it systematically degrades long-con

model-releasesarxiv-cs-cl
10 Jun 2026
Local Ai

Beyond Absolute Imitation: Anchored Residual Guidance for Privileged On-Policy Distillation

DGX agent

arXiv:2606.10385v1 Announce Type: cross Abstract: On-policy distillation (OPD) has demonstrated strong empirical gains in enhancing complex reasoning in LLMs by aligning a student model with a teacher

local-aiarxiv-cs-ai
10 Jun 2026
Model Releases

CleanPatrick: A Benchmark for Image Data Cleaning

DGX agent

arXiv:2505.11034v2 Announce Type: replace-cross Abstract: Robust machine learning depends on clean data, yet current image data cleaning benchmarks rely on synthetic noise or narrow human studies, lim

model-releasesarxiv-cs-ai
10 Jun 2026
Safety

Cross-Modal Knowledge Distillation without Paired Data: Theoretical Foundation and Algorithm

DGX agent

arXiv:2606.10504v1 Announce Type: new Abstract: Cross-modal knowledge distillation (CMKD) studies how a (large) teacher model trained on one type of data (e.g., images) can guide a (smaller) student m

safetyarxiv-cs-ai
10 Jun 2026
Model Releases

Cyst-X: A Multi-Center MRI Benchmark and Federated Learning Framework for Malignancy-Risk Stratification of Pancreatic Cystic Neoplasm

DGX agent

arXiv:2507.22017v4 Announce Type: replace-cross Abstract: Pancreatic cancer is projected to be the second-deadliest cancer by 2030, making early detection critical. Intraductal papillary mucinous neop

model-releasesarxiv-cs-cv
10 Jun 2026
Model Releases

Dario Amodei says he doesn't know what role Claude played in a missile strike on an Iranian school, and its use in this instance didn't violate Anthropic's ToS (Bloomberg)

DGX agent

Bloomberg: Dario Amodei says he doesn't know what role Claude played in a missile strike on an Iranian school, and its use in this instance didn't violate Anthropic's ToS — Anthropic PBC's boss said h

model-releasestechmeme
10 Jun 2026
Model Releases

Data-Driven Runway and Taxiway Exits Prediction of Landing Aircraft: A Case Study at Hartsfield-Jackson Atlanta International Airport

DGX agent

arXiv:2606.11017v1 Announce Type: new Abstract: Airport surface operations increasingly constrain performance at high-throughput hubs. This study examines arrival taxi-in decisions at Hartsfield-Jacks

model-releasesarxiv-cs-lg
10 Jun 2026
← Previous
1…686687688689690…1359
Next →