AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,577 results
26 Jun 2026

ProvenAI: Provenance-Native Traces of Evidence in Generated Answers

Model ReleasesDGX agent

arXiv:2606.26449v1 Announce Type: cross Abstract: Retrieval-augmented systems routinely present citations alongside generated answers, yet a citation does not confirm that the corresponding source mea

Quoting OpenAI

Model ReleasesDGX agent

We're beginning a limited preview of the GPT‑5.6 series: Sol, our flagship model; Terra, a balanced model for everyday work; and Luna, a fast and affordable model. Terra has competitive performance to

Qwen-Image-Agent: Bridging the Context Gap in Real-World Image Generation

Model ReleasesDGX agent

arXiv:2606.26907v1 Announce Type: new Abstract: While text-to-image (T2I) models have achieved remarkable progress, they struggle with real-world requests that are often underspecified, implicit, or d


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

R2D-RL: A RoboCup 2D Soccer Environment for Multi-Agent Reinforcement Learning

Model ReleasesDGX agent

arXiv:2606.18786v2 Announce Type: replace Abstract: Robot soccer is a challenging testbed for multi-agent reinforcement learning because it combines partial observability, cooperative and adversarial

Real-Time Safety Evaluation of Human Arm Operations Using a Wrist-Mounted IMU with PSM System

Model ReleasesDGX agent

arXiv:2502.09241v2 Announce Type: replace Abstract: This paper presents a novel approach to real-time safety monitoring in human-robot collaborative manufacturing environments through a wrist-mounted

ReasonCLIP-58M: Visually Grounded Commonsense Reasoning Supervision for CLIP

Model ReleasesDGX agent

arXiv:2606.26794v1 Announce Type: cross Abstract: CLIP and its variants are widely adopted visual backbones in multimodal systems, but their pretraining remains dominated by descriptive image-text ali

RedVox: Safety and Fairness Gaps in Speech Models Across Languages

Model ReleasesDGX agent

arXiv:2606.26968v1 Announce Type: new Abstract: Speech-capable models are increasingly deployed in real-world applications across languages. Yet their safety and fairness beyond English settings and u

Refusal Lives Downstream of Persona in Chat Models

Model ReleasesDGX agent

arXiv:2606.26161v1 Announce Type: new Abstract: Linear directions in activation space have been identified for both refusal and persona traits in instruction-tuned chat models, but the two have been s

Reinforcement Fine-Tuning of Flow-Matching Policies for Vision-Language-Action Models

Model ReleasesDGX agent

arXiv:2510.09976v2 Announce Type: replace Abstract: Vision-Language-Action (VLA) models such as OpenVLA, Octo, and pi_0 have shown strong generalization by leveraging large-scale demonstrations, yet t

Reinforcement Learning without Ground-Truth Solutions can Improve LLMs

Model ReleasesDGX agent

arXiv:2606.27369v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) for training LLMs typically rely on ground-truth answers to assign rewards, limiting their applica

ReportLogic: Evaluating Logical Quality in Deep Research Reports

Model ReleasesDGX agent

arXiv:2602.18446v2 Announce Type: replace-cross Abstract: Users increasingly rely on Large Language Models (LLMs) for Deep Research, using them to synthesize diverse sources into structured reports th

Representation Costs in Data Science: Foundations and the Quasi-Banach Spaces of Deep Neural Networks

Model ReleasesDGX agent

arXiv:2606.14954v3 Announce Type: replace-cross Abstract: We develop a general framework for analyzing representation costs of parametric data-fitting methods through their parameter-space regularizer

Reproducibility Study of 'AlphaEdit: Null-Space Constrained Knowledge Editing for Language Models'

Model ReleasesDGX agent

arXiv:2606.26783v1 Announce Type: cross Abstract: Fang et al. (2025) introduced a null-space constrained projection, named AlphaEdit, for locate-then-edit knowledge editing methods, theoretically guar

Ribbon: Scalable Approximation and Robust Uncertainty Quantification

Model ReleasesDGX agent

arXiv:2606.27269v1 Announce Type: cross Abstract: Reliably quantifying predictive uncertainty is difficult for complex, high-dimensional, or misspecified models. Both fully Bayesian and bootstrap resa

Rolling Shutter Relative Pose Estimation Made Practical

Model ReleasesDGX agent

arXiv:2606.26863v1 Announce Type: new Abstract: Rolling shutter (RS) cameras equip virtually all consumer devices, yet RS-aware relative pose estimation has remained impractical: the state-of-the-art

RoPEMover: Depth-Aware Object Relocation via Positional Embeddings

Model ReleasesDGX agent

arXiv:2606.27332v1 Announce Type: new Abstract: Moving an object in a single image requires geometry-consistent spatial rearrangement, including handling occlusions, revealing previously unseen region

RSPC: A Benchmark for Modeling Stress and Psychiatric Conditions in Digitally Mediated Relationships using Psychiatrist Annotations

Model ReleasesDGX agent

arXiv:2606.27247v1 Announce Type: new Abstract: In NLP, mental health conditions are often modeled as isolated phenomena, without interpersonal context. We use Reddit posts about long-distance relatio

Running the Gauntlet: Re-evaluating the Capabilities of Agents Beyond Familiar Environments

Model ReleasesDGX agent

arXiv:2606.14397v2 Announce Type: replace Abstract: As agentic systems continue to evolve and are widely deployed in real-world scenarios, there is a growing demand to faithfully evaluate their capabi

SamaVaani: Auditing and Debiasing Multilingual Clinical ASR for Indian Languages

Model ReleasesDGX agent

arXiv:2606.26901v1 Announce Type: cross Abstract: Automatic Speech Recognition (ASR) is increasingly used to document clinical encounters, yet its reliability in multilingual and demographically diver

Scalable AI-assisted Workflow Management for Detector Design Optimization Using Distributed Computing

Model ReleasesDGX agent

arXiv:2603.30014v2 Announce Type: replace-cross Abstract: The Production and Distributed Analysis (PanDA) system, originally developed for the ATLAS experiment at the CERN Large Hadron Collider (LHC),

Scaling Multi-Reference Image Generation with Dynamic Reward Optimization

Model ReleasesDGX agent

arXiv:2606.26947v1 Announce Type: cross Abstract: While personalized image generation has achieved remarkable progress, multi-reference image generation (MRIG) remains a challenging task. Most existin

scBench-Long: Verifiable Benchmarking of Long-Horizon Single-Cell Biology

Model ReleasesDGX agent

arXiv:2606.26563v1 Announce Type: cross Abstract: Single-cell studies require analysts to convert raw measurements into specific biological claims through multi-step workflows and integration of metad

SciFig: Towards Automating Editable Figure Generation for Scientific Papers

Model ReleasesDGX agent

arXiv:2601.04390v2 Announce Type: replace Abstract: High-quality methodology figures are central to scientific communication, yet they remain difficult and time-consuming to create. Such figures must

Securing agentic AI with perimeter guardrails: What's new in VPC Service Controls

Model ReleasesDGX agent

As enterprises scale autonomous AI agents into production, enabling safe innovation requires robust architectural guardrails. AI agents connect across tools and datasets, so it’s essential to establis

See & Sniff: Learning Visuo-Olfactory Representations

Model ReleasesDGX agent

arXiv:2606.27307v1 Announce Type: new Abstract: While modern multimodal models integrate vision with language, audio, or touch, olfaction remains largely unexplored due to the lack of paired visuo-olf

ShareLock: A Stealthy Multi-Tool Threshold Poisoning Attack Against MCP

Model ReleasesDGX agent

arXiv:2606.27027v1 Announce Type: cross Abstract: With the rapid evolution of LLM-driven agents, Model Context Protocol (MCP), an open protocol bridging LLMs with external tools, has quickly become fo

SharQ: Bridging Activation Sparsity and FP4 Quantization for LLM Inference

Model ReleasesDGX agent

arXiv:2606.26587v1 Announce Type: cross Abstract: Low-bit floating-point formats and semi-structured sparsity are increasingly supported by modern accelerators, yet combining them for LLM activation c

Signature filtering: a lightweight enhancement for statistical watermark detection in large language models

Model ReleasesDGX agent

arXiv:2606.18430v2 Announce Type: replace Abstract: Statistical watermarks help organizations attribute large language model (LLM) outputs, yet existing detectors often struggle when watermark signals

Simulation-based inference for rapid Bayesian parameter estimation in epidemiological models: a comparison with MCMC

Model ReleasesDGX agent

arXiv:2606.27286v1 Announce Type: new Abstract: Mechanistic epidemiological models are widely used to support infectious disease forecasting and public-health decision making. Bayesian calibration of

SITUATION EXPLAINED: Anthropic accused Alibaba of conducting a 'distillation attack' on Claude. We asked @ClementDelangue his thoughts: 'Dis…

Model ReleasesDGX agent

SITUATION EXPLAINED: Anthropic accused Alibaba of conducting a 'distillation attack' on Claude. We asked @ClementDelangue his thoughts: 'Distillation is a very common practice that everyone is using.

SITUATION EXPLAINED: What is at stake in the open source AI race? We asked @perrymetzger, chairman of Alliance for the Future. 'Open source …

Model ReleasesDGX agent

SITUATION EXPLAINED: What is at stake in the open source AI race? We asked @perrymetzger, chairman of Alliance for the Future. 'Open source models are a sort of cultural ambassador and civilizational

SocialPersona: Benchmarking Personalized Profiling and Response with Multimodal Social-Media Context

Model ReleasesDGX agent

arXiv:2606.26654v1 Announce Type: new Abstract: Personalized language-model assistants are often evaluated through a memory lens: can a model recall preferences users have explicitly stated in dialogu

Sources: DeepSeek's $7.4B raise was prompted by the release of Mythos as CEO Liang Wenfeng realized DeepSeek couldn't compete without a massive war chest (The Information)

Model ReleasesDGX agent

The Information: Sources: DeepSeek's $7.4B raise was prompted by the release of Mythos as CEO Liang Wenfeng realized DeepSeek couldn't compete without a massive war chest — Up until two months ago, De

Spurious Rewards Paradox: Mechanistically Understanding How RLVR Activates Memorization Shortcuts in LLMs

Model ReleasesDGX agent

arXiv:2601.11061v2 Announce Type: replace-cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) is highly effective for enhancing LLM reasoning, yet recent evidence shows models like Q

SSI-Policy: Learning Structured Scene Interfaces for Vision-Language Robotic Manipulation

Model ReleasesDGX agent

arXiv:2606.26800v1 Announce Type: new Abstract: Real-world robotic manipulation demands spatial grounding, task-aware reasoning, and precise control. Learning such capabilities becomes particularly ch

SSM Adapters via Hankel Reduced-order Modeling: Injection Site Determines Task Suitability in Long-Context Fine-Tuning

Model ReleasesDGX agent

arXiv:2606.26290v1 Announce Type: cross Abstract: While parameter-efficient fine-tuning (PEFT) typically targets attention projectors, its efficacy for tasks requiring sequential state accumulation re

Stochastic Gradient Optimization with Model-Assisted Sampling

Model ReleasesDGX agent

arXiv:2606.27171v1 Announce Type: new Abstract: This work addresses the problem of variance in stochastic gradient estimation for machine learning optimization. Deep learning relies on mini-batch meth

Taps sign

Model ReleasesDGX agent

Taps sign We believe in broad access and plan to make GPT-5.6 Sol, Terra, and Luna generally available in the coming weeks. For now, at the request of the U.S. government, we’re starting with a limite

Temporally Consistent Label Interpolation for Robust Surgical Multi-Task Learning under Challenging Conditions

Model ReleasesDGX agent

arXiv:2606.26634v1 Announce Type: new Abstract: Effective multi-task learning for surgical scene understanding is fundamentally hindered by annotation granularity mismatch; temporal workflow tasks suc

Term-Centric Hierarchy Induction from Heterogeneous Corpora

Model ReleasesDGX agent

arXiv:2606.26963v1 Announce Type: new Abstract: Organizing knowledge from diverse text sources into interpretable hierarchies is crucial for tasks such as policy analysis, innovation monitoring, and e

TF-TI2I: Training-Free Text-and-Image-to-Image Generation via Multi-Modal Implicit-Context Learning in Text-to-Image Models

Model ReleasesDGX agent

arXiv:2503.15283v2 Announce Type: replace Abstract: Text-and-Image-To-Image (TI2I), an extension of Text-To-Image (T2I), integrates image inputs with textual instructions to enhance image generation.

The Capability Frontier: Benchmarks Miss 82% of Model Performance

Model ReleasesDGX agent

arXiv:2606.26836v1 Announce Type: new Abstract: Existing benchmarks typically report accuracy for a single model on a single run. This systematically understates real-world LLM capabilities, particula

The Geometry of Updates: Fisher Alignment at Vocabulary Scale

Model ReleasesDGX agent

arXiv:2606.27242v1 Announce Type: cross Abstract: Training-free source selection for LLM families with shared vocabularies arises in scientific string domains such as SMILES, protein, and genomic sequ

The Inattentional Gap: Task-Conditioned Language and Vision Models Omit the Safety-Critical Signals They Can Otherwise Report

Model ReleasesDGX agent

arXiv:2606.26529v1 Announce Type: cross Abstract: AI safety is evaluated by how reliably a model detects the hazards it is told to find, yet accidents often arise from the hazard no one specified. We

The Open Source Economic Index of AI Adoption and Capability

Model ReleasesDGX agent

arXiv:2606.26118v1 Announce Type: cross Abstract: We work towards measuring both AI adoption and the capability of AI to perform discrete labor tasks across various occupations. To measure adoption, w

The Red Queen Godel Machine: Co-Evolving Agents and Their Evaluators

Model ReleasesDGX agent

arXiv:2606.26294v1 Announce Type: cross Abstract: Self-improving agents are state-of-the-art (SOTA) on agentic coding benchmarks and have recently been extended to general domains. However, their sear

The Spec Growth Engine: Spec-Anchored, Code-Coupled, Drift-Enforced Architecture for AI-Assisted Software Development

Model ReleasesDGX agent

arXiv:2606.27045v1 Announce Type: cross Abstract: AI coding agents dramatically accelerate implementation speed but introduce two structural failure modes that existing spec-driven approaches do not f

The strongest models are gated and access is granted only to a select few. Hermes Agent now exposes MoA presets as virtual models, giving yo…

Model ReleasesDGX agent

The strongest models are gated and access is granted only to a select few. Hermes Agent now exposes MoA presets as virtual models, giving you capabilities beyond the publicly available frontier: 8% hi

Theory-Scale Auto-Formalization of Logics for Computer Science

Model ReleasesDGX agent

arXiv:2606.26525v1 Announce Type: new Abstract: Auto-formalization is critical for scalable formal verification, but existing progress largely focuses on isolated statements, while theory-scale auto-f

Thinking Like a Scientist? A Structural Study of LLM-Generated Research Methods

Model ReleasesDGX agent

arXiv:2606.26130v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used to guide research methodology, yet their default methodological tendencies under minimal prompting

This week someone targeted my family for harm with a false report. We’re physically OK, but that doesn’t mean we weren’t harmed. I am beyond…

Model ReleasesDGX agent

This week someone targeted my family for harm with a false report. We’re physically OK, but that doesn’t mean we weren’t harmed. I am beyond furious. Whatever your politics, this is awful, wrong, and

TMP: Tree-structured Mixed-policy Pruning for Large-scale Image Generation and Editing

Model ReleasesDGX agent

arXiv:2606.27089v1 Announce Type: new Abstract: Modern image generation model rapidly grows their sizes to meet high-fidelity image synthesis. However, they gradually become unaffordable for their eno

Today on TITV: -Trump asks OpenAI to stagger release of new model | @leomschwartz & @amir, The Information -Google pressures publishers on A…

Model ReleasesDGX agent

Today on TITV: -Trump asks OpenAI to stagger release of new model | @leomschwartz & @amir, The Information -Google pressures publishers on AI licensing | @anngehan -Inside an AI power user’s agent wor

Towards Explainable Adjudicative Variance: Quantifying Judicial Discretion via Gated Multi-Task Learning

Model ReleasesDGX agent

arXiv:2606.27069v1 Announce Type: new Abstract: Legal outcome prediction must disentangle objective case facts from adjudicative context. Merit-based rulings rely on factual evidence while technical d

Towards Video Anomaly Detection from Event Streams: A Baseline and Benchmark Datasets

Model ReleasesDGX agent

arXiv:2603.24991v2 Announce Type: replace Abstract: Event-based vision, characterized by low redundancy, focus on dynamic motion, and inherent privacy-preserving properties, naturally fits the demands

TraMP-LLaMA: Generative Interpretability with Decoupled Instruction Tuning for Facial Expression Quality Assessment

Model ReleasesDGX agent

arXiv:2606.26942v1 Announce Type: new Abstract: Existing facial expression quality assessment (FEQA) methods typically produce only a severity score, without explicitly communicating the observable fa

Tuning Language Models by Mixture-of-Depths Ensemble

Model ReleasesDGX agent

arXiv:2410.13077v2 Announce Type: replace-cross Abstract: Transformer-based Large Language Models (LLMs) traditionally rely on final-layer loss for finetuning and final-layer representations for predi

UAV-MapFusion: RTK-Aligned Uncertainty-Aware Coarse-to-Fine Multi-Session UAV Mapping

Model ReleasesDGX agent

arXiv:2606.26928v1 Announce Type: new Abstract: Large-scale point cloud maps are essential for robotics and spatial intelligence tasks. UAVs provide an efficient means for large-scale map acquisition;

UBS says 60% of companies now watching AI budgets are moving to cheaper models and open-source Chinese models The pressure is coming from ex…

Model ReleasesDGX agent

UBS says 60% of companies now watching AI budgets are moving to cheaper models and open-source Chinese models The pressure is coming from extreme bills, including users spending up to $35K/month, team

Unconventional AI debuts oscillator-based Un-0 model series

Model ReleasesDGX agent

Unconventional AI Inc. has developed an artificial intelligence architecture that could improve the power efficiency of image generation models. The technology is the basis of a new neural network ser

← Previous
1…129130131132133…377
Next →