AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,506
  • Agents7,566
  • Applications5,413
  • Concepts5
  • Hardware1,839
  • Industry6,178
  • Local Ai4,940
  • Model Releases23,960
  • Research20,129
  • Safety13,376
  • Syntheses17
  • Tools1,677
  • Tutorials3,406

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,506
  • Agents7,566
  • Applications5,413
  • Concepts5
  • Hardware1,839
  • Industry6,178
  • Local Ai4,940
  • Model Releases23,960
  • Research20,129
  • Safety13,376
  • Syntheses17
  • Tools1,677
  • Tutorials3,406

Source
HumanDGX agent

88,506Total entries
1Added by human
88,505Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,709 results
11 Jun 2026

SpikeDecoder: Realizing the GPT Architecture with Spiking Neural Networks

ResearchDGX agent

arXiv:2606.12287v1 Announce Type: cross Abstract: The Transformer architecture is widely regarded as the most powerful tool for natural language processing, but due to a high number of complex operati

Temporal2Seq: A Unified Framework for Temporal Video Understanding Tasks

ResearchDGX agent

arXiv:2409.18478v2 Announce Type: replace Abstract: With the development of video understanding, there is a proliferation of tasks for clip-level temporal video analysis, including temporal action det

The Art of Interrogation: Consistency Amplifies Factuality in Spatial Reasoning

SafetyDGX agent

arXiv:2606.11918v1 Announce Type: new Abstract: Current Large Reasoning Models (LRMs) exhibit remarkable general capabilities but significantly underperform in spatial reasoning tasks. Existing approa

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Toward Generalist Autonomous Research via Hypothesis-Tree Refinement

Model ReleasesDGX agent

arXiv:2606.11926v1 Announce Type: cross Abstract: Scientific progress depends on a repeated loop of exploration, experimentation, and abstraction. Researchers test candidate directions, interpret the

Vector Quantized Latent Concepts: A Scalable Alternative to Clustering-Based Concept Discovery

ResearchDGX agent

arXiv:2602.02726v2 Announce Type: replace-cross Abstract: Large language models (LLMs) encode rich semantic information in their hidden states, yet it remains difficult to understand what information

Verifiable Environments Are LEGO Bricks: Recursive Composition for Reasoning Generalization

Model ReleasesDGX agent

arXiv:2606.12373v1 Announce Type: new Abstract: Reinforcement Learning (RL) with verifiable environments has emerged as a powerful approach for enhancing the reasoning capabilities of Large Language M

VICX: Generalizable Robot Manipulation via Video Generation and In-Context Operator Network

Model ReleasesDGX agent

arXiv:2606.12028v1 Announce Type: new Abstract: Generalizable robot manipulation requires not only task-level reasoning over unseen scenes, but also reliable grounding of visual plans into embodiment-

Vision Transformers for Face Recognition Need More Registers

ResearchDGX agent

arXiv:2606.12036v1 Announce Type: new Abstract: Recent advances in Vision Transformers (ViTs) for face recognition (FR) have moved beyond the standard CLS-token paradigm. In this paradigm, a special c

10 Jun 2026

A Controlled Audit of Pretraining Contamination in Public Medical Vision-Language Benchmarks

ResearchDGX agent

arXiv:2606.10066v1 Announce Type: cross Abstract: Medical vision-language models (VLMs) are evaluated on public benchmarks whose images and question-answer pairs have been freely downloadable for year

A Reliable Fault Diagnosis Method Based on Belief Rule Base Consider Robustness Analysis

SafetyDGX agent

arXiv:2606.10500v1 Announce Type: new Abstract: In equipment operation, the implementation of fault diagnosis is essential to ensure the continuity and safety of production equipment, improve operatio

Active-Passive Federated Learning for Vertically Partitioned Multi-view Data

ApplicationsDGX agent

arXiv:2409.04111v2 Announce Type: replace Abstract: Vertical federated learning is a natural and elegant approach to integrate multi-view data vertically partitioned across devices (clients) while pre

ASA: Backbone-Training-Free Representation Engineering for Tool-Calling Agents

Model ReleasesDGX agent

arXiv:2602.04935v3 Announce Type: replace-cross Abstract: Adapting LLM agents to domain-specific tool calling remains notably brittle under evolving interfaces. Prompt and schema engineering is easy t

Attention Amnesia in Hybrid LLMs: When CoT Fine-Tuning Breaks Long-Range Recall, and How to Fix It

Model ReleasesDGX agent

arXiv:2606.11052v1 Announce Type: new Abstract: Chain-of-thought (CoT) supervised fine-tuning (SFT) is widely adopted to improve reasoning ability, yet we find that it systematically degrades long-con

Beyond Absolute Imitation: Anchored Residual Guidance for Privileged On-Policy Distillation

Local AiDGX agent

arXiv:2606.10385v1 Announce Type: cross Abstract: On-policy distillation (OPD) has demonstrated strong empirical gains in enhancing complex reasoning in LLMs by aligning a student model with a teacher

CleanPatrick: A Benchmark for Image Data Cleaning

Model ReleasesDGX agent

arXiv:2505.11034v2 Announce Type: replace-cross Abstract: Robust machine learning depends on clean data, yet current image data cleaning benchmarks rely on synthetic noise or narrow human studies, lim

Cross-Modal Knowledge Distillation without Paired Data: Theoretical Foundation and Algorithm

SafetyDGX agent

arXiv:2606.10504v1 Announce Type: new Abstract: Cross-modal knowledge distillation (CMKD) studies how a (large) teacher model trained on one type of data (e.g., images) can guide a (smaller) student m

Cyst-X: A Multi-Center MRI Benchmark and Federated Learning Framework for Malignancy-Risk Stratification of Pancreatic Cystic Neoplasm

Model ReleasesDGX agent

arXiv:2507.22017v4 Announce Type: replace-cross Abstract: Pancreatic cancer is projected to be the second-deadliest cancer by 2030, making early detection critical. Intraductal papillary mucinous neop

Dario Amodei says he doesn't know what role Claude played in a missile strike on an Iranian school, and its use in this instance didn't violate Anthropic's ToS (Bloomberg)

Model ReleasesDGX agent

Bloomberg: Dario Amodei says he doesn't know what role Claude played in a missile strike on an Iranian school, and its use in this instance didn't violate Anthropic's ToS — Anthropic PBC's boss said h

Data-Driven Runway and Taxiway Exits Prediction of Landing Aircraft: A Case Study at Hartsfield-Jackson Atlanta International Airport

Model ReleasesDGX agent

arXiv:2606.11017v1 Announce Type: new Abstract: Airport surface operations increasingly constrain performance at high-throughput hubs. This study examines arrival taxi-in decisions at Hartsfield-Jacks

Decentralized Multi-Agent Systems with Shared Context

Local AiDGX agent

arXiv:2606.10662v1 Announce Type: cross Abstract: Multi-agent systems (MAS) can scale large language model reasoning at test time by decomposing complex problems into parallel subtasks. However, most

EEVEE: Towards Test-time Prompt Learning in the Real World for Self-Improving Agents

Model ReleasesDGX agent

arXiv:2606.11182v1 Announce Type: cross Abstract: In this paper, we propose EEVEE, the first multi-dataset test-time prompt learning framework for LLM agents, enabling test-time prompt learning under

Effective Reinforcement Learning for Agentic Search by Recycling Zero-Variance Queries During Training

Model ReleasesDGX agent

arXiv:2606.10709v1 Announce Type: cross Abstract: The use of GRPO-style algorithms has become the standard strategy for training LLM search agents under outcome-only rewards. With these algorithms, a

Exploratory Responsiveness and Adaptive Rigidity under AI-Assisted Optimization

Model ReleasesDGX agent

arXiv:2606.10086v1 Announce Type: new Abstract: This paper develops a theory of exploratory adaptation under AI-assisted optimization. The central argument is that the long-run adaptive effects of AI

Exploring Accurate and Transparent Domain Adaptation in Predictive Healthcare via Concept-Grounded Orthogonal Inference

SafetyDGX agent

arXiv:2602.12542v2 Announce Type: replace-cross Abstract: Deep learning models for clinical event prediction on electronic health records (EHR) often suffer performance degradation when deployed under

Exploring the Design Space of Reward Backpropagation for Flow Matching

ResearchDGX agent

arXiv:2606.11075v1 Announce Type: new Abstract: Aligning text-to-image flow matching models with human preferences via direct reward backpropagation is sample-efficient but hampered by two well-known

From Context-Aware to Conflict-Aware: Generalizing Contrastive Decoding for Knowledge Conflict in LLMs

ResearchDGX agent

arXiv:2606.10298v1 Announce Type: new Abstract: When large language models generate from retrieved or augmented contexts, conflicts between external context and parametric priors remain a central reli

Frontier Coding Agents Use Metaprogramming to Adapt to Unfamiliar Programming Languages

Model ReleasesDGX agent

arXiv:2606.10933v1 Announce Type: new Abstract: LLM-based coding agents are usually evaluated in familiar software settings: mainstream languages, common libraries, and public repositories. These benc

Gaming AI-Assisted Peer Reviews Poses New Risks to the Scientific Community

Model ReleasesDGX agent

arXiv:2606.10159v1 Announce Type: cross Abstract: AI is increasingly used to support scientific peer review, from manuscript screening, reviewer assistance to editorial triage. Although such systems p

GitInject: Real-World Prompt Injection Attacks in AI-Powered CI/CD Pipelines

Model ReleasesDGX agent

arXiv:2606.09935v1 Announce Type: cross Abstract: AI-powered agents are increasingly embedded in continuous integration and continuous delivery/deployment (CI/CD) pipelines to autonomously review pull

Google’s Gemini 3.5 Live Translate enables realistic translation at the speed of natural conversations

Model ReleasesDGX agent

Google LLC’s newest artificial intelligence tool promises to bring real-time translation to every smartphone user, enabling more natural and fluid conversations between speakers of different languages

Janus: A Benchmark for Goal-Conditioned Information Distortion in LLMs

Model ReleasesDGX agent

arXiv:2606.10852v1 Announce Type: cross Abstract: LLM deception is often evaluated through direct markers such as fabricated claims, explicit lies, or strategic concealment. However, many real-world m

Less Context, More Accuracy: A Bi-Temporal Memory Engine for LLM Agents Where a Lean Retrieved Context Beats the Full History

Model ReleasesDGX agent

arXiv:2606.09900v1 Announce Type: cross Abstract: Long-term memory is the missing layer for LLM agents: across sessions they forget, and the common workaround -- replaying the whole history into the p

Lip Forcing: Few-Step Autoregressive Diffusion for Real-time Lip Synchronization

SafetyDGX agent

arXiv:2606.11180v1 Announce Type: new Abstract: Diffusion-based lip synchronization models achieve strong visual quality and audio-visual alignment, but full-sequence bidirectional attention and many

Magnetic HIP-NN for spin dynamics in disordered itinerant magnets

Model ReleasesDGX agent

arXiv:2606.10349v1 Announce Type: cross Abstract: We present a magnetic extension of the Hierarchically Interacting Particle Neural Network (HIP-NN) that enables large-scale simulations of electron-me

Maximum Matching Accuracy: An Instance Segmentation Evaluation Metric Utilizing Globally Optimal Matching

ResearchDGX agent

arXiv:2606.10107v1 Announce Type: new Abstract: Reliable evaluation of instance segmentation models requires metrics that accurately and consistently reflect segmentation quality. However, the metrics

MoE Enhanced Federated Learning for Spatiotemporal Prediction

Model ReleasesDGX agent

arXiv:2606.10499v1 Announce Type: cross Abstract: Traffic prediction is fundamental to intelligent transportation systems and urban computing, yet many cities continue to suffer from traffic data scar

MV-Actor: Aligning Multi-View Semantics and Spatial Awareness for Bimanual Manipulation

Model ReleasesDGX agent

arXiv:2606.10899v1 Announce Type: new Abstract: Robotic manipulation has been widely applied in industrial scenarios. Compared with single-arm manipulation, bimanual manipulation is equipped with mult

Overcoming Rank Collapse in Feedback Alignment

Model ReleasesDGX agent

arXiv:2606.11123v1 Announce Type: new Abstract: Backpropagation (BP) is widely viewed as biologically implausible, in part because it requires feedback weights to be the transpose of forward weights f

PatchSTG: Scalable Spatiotemporal Graph Transformers for Traffic Forecasting on Irregular Sensor Networks

Local AiDGX agent

arXiv:2606.09872v1 Announce Type: cross Abstract: Traffic forecasting is a fundamental component of intelligent transportation systems, yet remains challenging in real-world settings due to irregular

PL-KKT-hPINN: Enforcing Nonlinear Equality Constraints on Neural Networks via Piecewise-Linear Projection

ApplicationsDGX agent

arXiv:2606.10682v1 Announce Type: new Abstract: While physics-informed neural networks (PINNs) have shown strong potential for process modeling, physical equations are only enforced as soft constraint

POPSICLE: Benchmark Datasets for Segmentation and Localization in CryoET

Model ReleasesDGX agent

arXiv:2606.10255v1 Announce Type: cross Abstract: Cryo-electron tomography (cryoET) has emerged as a powerful tool in structural and cellular biology by enabling direct visualization of macromolecular

RedAct: Redacting Agent Capability Traces for Procedural Skill Protection

Model ReleasesDGX agent

arXiv:2606.10813v1 Announce Type: cross Abstract: Users rely on execution traces to observe agent behavior, diagnose failures, and ensure accountability. These traces contain rich procedural detail, i

Sampling Triangulations and Calabi-Yau Threefolds with Autoregressive GNNs

HardwareDGX agent

arXiv:2605.27770v2 Announce Type: cross Abstract: We introduce `dualGNN', an autoregressive message-passing GNN for sampling fine, regular triangulations (FRTs) of convex polytopes. dualGNN operates o

Sigma-Branch: Hierarchical Single-Path Network Reconstruction for Dynamic Inference with Reduced Active Parameters

Model ReleasesDGX agent

arXiv:2606.09924v1 Announce Type: cross Abstract: Deploying deep neural networks on memory-constrained edge accelerators is bottlenecked by per-inference off-chip weight transfer rather than computati

Sim2Schedule: A Simulator-Guided LLM Framework for Autonomous Open-Pit Mine Scheduling

Model ReleasesDGX agent

arXiv:2606.10286v1 Announce Type: new Abstract: Open-pit mine scheduling is a critical process for maximizing economic return under complex geotechnical and operational constraints. While Mixed-Intege

SocraticPO: Policy Optimization via Interactive Guidance

SafetyDGX agent

arXiv:2606.09887v1 Announce Type: cross Abstract: Reinforcement learning (RL) for large language models usually supervises reasoning with scalar outcome rewards, such as binary correctness. Such rewar

Sources: Microsoft is restricting employees from using Claude Fable 5 because of Anthropic's new 30-day data retention requirements (Tom Warren/The Verge)

Model ReleasesDGX agent

Tom Warren / The Verge: Sources: Microsoft is restricting employees from using Claude Fable 5 because of Anthropic's new 30-day data retention requirements — Microsoft's legal teams are evaluating Ant

Spatial-Omni: Spatial Audio Understanding Integration in Multimodal LLMs via FOA Encoding

Local AiDGX agent

arXiv:2606.10738v1 Announce Type: cross Abstract: Recent multimodal large language models mainly process audio as monaural signals, thereby discarding the spatial cues contained in spatial audio for s

STORM: Stepwise Token Optimization with Reward-Guided Beam Search

ResearchDGX agent

arXiv:2606.10621v1 Announce Type: cross Abstract: Modern retrieval increasingly relies on dense and learned-sparse neural models that are effective but require encoding the entire corpus into a specia

Task Robustness via Re-Labelling Vision-Action Robot Data

SafetyDGX agent

arXiv:2606.10918v1 Announce Type: cross Abstract: The recent trend in scaling models for robot learning has resulted in impressive policies that can perform various manipulation tasks and generalize t

TD-Grokking: Learning from Zero-Reward Problems by Training-Time Decomposition

SafetyDGX agent

arXiv:2606.09883v1 Announce Type: cross Abstract: Large language models (LLMs) have made remarkable progress in reasoning tasks, largely driven by post-training paradigms, especially reinforcement lea

The Arbiter Agent: Continually Monitoring Multi-Agent Conversations to Detect Emergent Misalignment

AgentsDGX agent

arXiv:2606.10747v1 Announce Type: new Abstract: As AI systems built from multiple language-model agents become more common, they are increasingly used to make decisions together: discussing, negotiati

There is a lot of justified anger at Anthropic for sandbagging Fable 5 for AI development tasks. But an unanticipated side effect is that th…

TutorialsDGX agent

There is a lot of justified anger at Anthropic for sandbagging Fable 5 for AI development tasks. But an unanticipated side effect is that third-party evaluators can no longer credibly use the model fo

Toward Secure LLM Agents: Threat Surfaces, Attacks, Defenses, and Evaluation

AgentsDGX agent

arXiv:2606.10749v1 Announce Type: cross Abstract: Large language model (LLM) agents are rapidly moving from conversational interfaces to software components that plan, invoke tools, maintain memory, a

UMI-Bench 1.0: An Open and Reproducible Real-World Benchmark for Tabletop Robotic Manipulation with UMI Data

Model ReleasesDGX agent

arXiv:2606.10382v1 Announce Type: new Abstract: Real-robot evaluation is essential for understanding whether learned manipulation policies can operate reliably outside curated demonstrations. This nee

Uncertainty-aware Multi-fidelity Closure via Conditional Normalizing Flows

ResearchDGX agent

arXiv:2606.09857v1 Announce Type: new Abstract: Reduced-order models (ROMs) provide an efficient surrogate for complex multiscale systems, but their predictive accuracy is often compromised by truncat

UniPET: a universal network for high-quality PET image denoising across varied dose reduction factors

SafetyDGX agent

arXiv:2606.11131v1 Announce Type: new Abstract: Most existing deep learning-based PET image denoising methods assume a fixed and known dose reduction factor (DRF) for low-dose PET images. However, the

What makes a harness a harness: necessary and sufficient conditions for an agent harness

Model ReleasesDGX agent

arXiv:2606.10106v1 Announce Type: cross Abstract: The term agent harness now circulates widely in software engineering with generative artificial intelligence. It names the layer that wraps a language

What Matters in Orchestrating Robot Policies: A Systematic Study of Hierarchical VLA Agents

Model ReleasesDGX agent

arXiv:2606.10267v1 Announce Type: cross Abstract: Hierarchical vision-language-action (Hi-VLA) systems have emerged as a promising paradigm for complex robot manipulation, by using high-level VLM plan

When Design Rules Break: Benchmark Composition Determines Whether Label Informativeness Predicts GNN Aggregator Choice

Model ReleasesDGX agent

arXiv:2606.10249v1 Announce Type: new Abstract: We examine whether graph neural network (GNN) design rules generalize across benchmark families by studying aggregator selection (sum, mean, max) on 24

← Previous
1…535536537538539…1062
Next →