AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,595 results
Model Releases

RedVox: Safety and Fairness Gaps in Speech Models Across Languages

DGX agent

arXiv:2606.26968v1 Announce Type: new Abstract: Speech-capable models are increasingly deployed in real-world applications across languages. Yet their safety and fairness beyond English settings and u

model-releasesarxiv-cs-cl
26 Jun 2026
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Refusal Lives Downstream of Persona in Chat Models

DGX agent

arXiv:2606.26161v1 Announce Type: new Abstract: Linear directions in activation space have been identified for both refusal and persona traits in instruction-tuned chat models, but the two have been s

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

Reinforcement Fine-Tuning of Flow-Matching Policies for Vision-Language-Action Models

DGX agent

arXiv:2510.09976v2 Announce Type: replace Abstract: Vision-Language-Action (VLA) models such as OpenVLA, Octo, and pi_0 have shown strong generalization by leveraging large-scale demonstrations, yet t

model-releasesarxiv-cs-lg
26 Jun 2026
Model Releases

Reinforcement Learning without Ground-Truth Solutions can Improve LLMs

DGX agent

arXiv:2606.27369v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) for training LLMs typically rely on ground-truth answers to assign rewards, limiting their applica

model-releasesarxiv-cs-lg
26 Jun 2026
Model Releases

ReportLogic: Evaluating Logical Quality in Deep Research Reports

DGX agent

arXiv:2602.18446v2 Announce Type: replace-cross Abstract: Users increasingly rely on Large Language Models (LLMs) for Deep Research, using them to synthesize diverse sources into structured reports th

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

Representation Costs in Data Science: Foundations and the Quasi-Banach Spaces of Deep Neural Networks

DGX agent

arXiv:2606.14954v3 Announce Type: replace-cross Abstract: We develop a general framework for analyzing representation costs of parametric data-fitting methods through their parameter-space regularizer

model-releasesarxiv-cs-lg
26 Jun 2026
Model Releases

Reproducibility Study of 'AlphaEdit: Null-Space Constrained Knowledge Editing for Language Models'

DGX agent

arXiv:2606.26783v1 Announce Type: cross Abstract: Fang et al. (2025) introduced a null-space constrained projection, named AlphaEdit, for locate-then-edit knowledge editing methods, theoretically guar

model-releasesarxiv-cs-cl
26 Jun 2026
Model Releases

Ribbon: Scalable Approximation and Robust Uncertainty Quantification

DGX agent

arXiv:2606.27269v1 Announce Type: cross Abstract: Reliably quantifying predictive uncertainty is difficult for complex, high-dimensional, or misspecified models. Both fully Bayesian and bootstrap resa

model-releasesarxiv-cs-lg
26 Jun 2026
Model Releases

Rolling Shutter Relative Pose Estimation Made Practical

DGX agent

arXiv:2606.26863v1 Announce Type: new Abstract: Rolling shutter (RS) cameras equip virtually all consumer devices, yet RS-aware relative pose estimation has remained impractical: the state-of-the-art

model-releasesarxiv-cs-cv
26 Jun 2026
Model Releases

RoPEMover: Depth-Aware Object Relocation via Positional Embeddings

DGX agent

arXiv:2606.27332v1 Announce Type: new Abstract: Moving an object in a single image requires geometry-consistent spatial rearrangement, including handling occlusions, revealing previously unseen region

model-releasesarxiv-cs-cv
26 Jun 2026
Model Releases

RSPC: A Benchmark for Modeling Stress and Psychiatric Conditions in Digitally Mediated Relationships using Psychiatrist Annotations

DGX agent

arXiv:2606.27247v1 Announce Type: new Abstract: In NLP, mental health conditions are often modeled as isolated phenomena, without interpersonal context. We use Reddit posts about long-distance relatio

model-releasesarxiv-cs-lg
26 Jun 2026
Model Releases

Running the Gauntlet: Re-evaluating the Capabilities of Agents Beyond Familiar Environments

DGX agent

arXiv:2606.14397v2 Announce Type: replace Abstract: As agentic systems continue to evolve and are widely deployed in real-world scenarios, there is a growing demand to faithfully evaluate their capabi

model-releasesarxiv-cs-lg
26 Jun 2026
Model Releases

SamaVaani: Auditing and Debiasing Multilingual Clinical ASR for Indian Languages

DGX agent

arXiv:2606.26901v1 Announce Type: cross Abstract: Automatic Speech Recognition (ASR) is increasingly used to document clinical encounters, yet its reliability in multilingual and demographically diver

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

Scalable AI-assisted Workflow Management for Detector Design Optimization Using Distributed Computing

DGX agent

arXiv:2603.30014v2 Announce Type: replace-cross Abstract: The Production and Distributed Analysis (PanDA) system, originally developed for the ATLAS experiment at the CERN Large Hadron Collider (LHC),

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

Scaling Multi-Reference Image Generation with Dynamic Reward Optimization

DGX agent

arXiv:2606.26947v1 Announce Type: cross Abstract: While personalized image generation has achieved remarkable progress, multi-reference image generation (MRIG) remains a challenging task. Most existin

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

scBench-Long: Verifiable Benchmarking of Long-Horizon Single-Cell Biology

DGX agent

arXiv:2606.26563v1 Announce Type: cross Abstract: Single-cell studies require analysts to convert raw measurements into specific biological claims through multi-step workflows and integration of metad

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

SciFig: Towards Automating Editable Figure Generation for Scientific Papers

DGX agent

arXiv:2601.04390v2 Announce Type: replace Abstract: High-quality methodology figures are central to scientific communication, yet they remain difficult and time-consuming to create. Such figures must

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

Securing agentic AI with perimeter guardrails: What's new in VPC Service Controls

DGX agent

As enterprises scale autonomous AI agents into production, enabling safe innovation requires robust architectural guardrails. AI agents connect across tools and datasets, so it’s essential to establis

model-releasesgoogle-cloud-ai
26 Jun 2026
Model Releases

See & Sniff: Learning Visuo-Olfactory Representations

DGX agent

arXiv:2606.27307v1 Announce Type: new Abstract: While modern multimodal models integrate vision with language, audio, or touch, olfaction remains largely unexplored due to the lack of paired visuo-olf

model-releasesarxiv-cs-cv
26 Jun 2026
Model Releases

ShareLock: A Stealthy Multi-Tool Threshold Poisoning Attack Against MCP

DGX agent

arXiv:2606.27027v1 Announce Type: cross Abstract: With the rapid evolution of LLM-driven agents, Model Context Protocol (MCP), an open protocol bridging LLMs with external tools, has quickly become fo

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

SharQ: Bridging Activation Sparsity and FP4 Quantization for LLM Inference

DGX agent

arXiv:2606.26587v1 Announce Type: cross Abstract: Low-bit floating-point formats and semi-structured sparsity are increasingly supported by modern accelerators, yet combining them for LLM activation c

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

Signature filtering: a lightweight enhancement for statistical watermark detection in large language models

DGX agent

arXiv:2606.18430v2 Announce Type: replace Abstract: Statistical watermarks help organizations attribute large language model (LLM) outputs, yet existing detectors often struggle when watermark signals

model-releasesarxiv-cs-lg
26 Jun 2026
Model Releases

Simulation-based inference for rapid Bayesian parameter estimation in epidemiological models: a comparison with MCMC

DGX agent

arXiv:2606.27286v1 Announce Type: new Abstract: Mechanistic epidemiological models are widely used to support infectious disease forecasting and public-health decision making. Bayesian calibration of

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

SITUATION EXPLAINED: Anthropic accused Alibaba of conducting a 'distillation attack' on Claude. We asked @ClementDelangue his thoughts: 'Dis…

DGX agent

SITUATION EXPLAINED: Anthropic accused Alibaba of conducting a 'distillation attack' on Claude. We asked @ClementDelangue his thoughts: 'Distillation is a very common practice that everyone is using.

model-releasesclem-delangue--x
26 Jun 2026
Model Releases

SITUATION EXPLAINED: What is at stake in the open source AI race? We asked @perrymetzger, chairman of Alliance for the Future. 'Open source …

DGX agent

SITUATION EXPLAINED: What is at stake in the open source AI race? We asked @perrymetzger, chairman of Alliance for the Future. 'Open source models are a sort of cultural ambassador and civilizational

model-releasesclem-delangue--x
26 Jun 2026
Model Releases

SocialPersona: Benchmarking Personalized Profiling and Response with Multimodal Social-Media Context

DGX agent

arXiv:2606.26654v1 Announce Type: new Abstract: Personalized language-model assistants are often evaluated through a memory lens: can a model recall preferences users have explicitly stated in dialogu

model-releasesarxiv-cs-cl
26 Jun 2026
Model Releases

Sources: DeepSeek's $7.4B raise was prompted by the release of Mythos as CEO Liang Wenfeng realized DeepSeek couldn't compete without a massive war chest (The Information)

DGX agent

The Information: Sources: DeepSeek's $7.4B raise was prompted by the release of Mythos as CEO Liang Wenfeng realized DeepSeek couldn't compete without a massive war chest — Up until two months ago, De

model-releasestechmeme
26 Jun 2026
Model Releases

Spurious Rewards Paradox: Mechanistically Understanding How RLVR Activates Memorization Shortcuts in LLMs

DGX agent

arXiv:2601.11061v2 Announce Type: replace-cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) is highly effective for enhancing LLM reasoning, yet recent evidence shows models like Q

model-releasesarxiv-cs-cl
26 Jun 2026
Model Releases

SSI-Policy: Learning Structured Scene Interfaces for Vision-Language Robotic Manipulation

DGX agent

arXiv:2606.26800v1 Announce Type: new Abstract: Real-world robotic manipulation demands spatial grounding, task-aware reasoning, and precise control. Learning such capabilities becomes particularly ch

model-releasesarxiv-cs-ro
26 Jun 2026
Model Releases

SSM Adapters via Hankel Reduced-order Modeling: Injection Site Determines Task Suitability in Long-Context Fine-Tuning

DGX agent

arXiv:2606.26290v1 Announce Type: cross Abstract: While parameter-efficient fine-tuning (PEFT) typically targets attention projectors, its efficacy for tasks requiring sequential state accumulation re

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

Stochastic Gradient Optimization with Model-Assisted Sampling

DGX agent

arXiv:2606.27171v1 Announce Type: new Abstract: This work addresses the problem of variance in stochastic gradient estimation for machine learning optimization. Deep learning relies on mini-batch meth

model-releasesarxiv-cs-lg
26 Jun 2026
Model Releases

Taps sign

DGX agent

Taps sign We believe in broad access and plan to make GPT-5.6 Sol, Terra, and Luna generally available in the coming weeks. For now, at the request of the U.S. government, we’re starting with a limite

model-releasescohere--x
26 Jun 2026
Model Releases

Temporally Consistent Label Interpolation for Robust Surgical Multi-Task Learning under Challenging Conditions

DGX agent

arXiv:2606.26634v1 Announce Type: new Abstract: Effective multi-task learning for surgical scene understanding is fundamentally hindered by annotation granularity mismatch; temporal workflow tasks suc

model-releasesarxiv-cs-cv
26 Jun 2026
Model Releases

Term-Centric Hierarchy Induction from Heterogeneous Corpora

DGX agent

arXiv:2606.26963v1 Announce Type: new Abstract: Organizing knowledge from diverse text sources into interpretable hierarchies is crucial for tasks such as policy analysis, innovation monitoring, and e

model-releasesarxiv-cs-cl
26 Jun 2026
Model Releases

TF-TI2I: Training-Free Text-and-Image-to-Image Generation via Multi-Modal Implicit-Context Learning in Text-to-Image Models

DGX agent

arXiv:2503.15283v2 Announce Type: replace Abstract: Text-and-Image-To-Image (TI2I), an extension of Text-To-Image (T2I), integrates image inputs with textual instructions to enhance image generation.

model-releasesarxiv-cs-cv
26 Jun 2026
Model Releases

The Capability Frontier: Benchmarks Miss 82% of Model Performance

DGX agent

arXiv:2606.26836v1 Announce Type: new Abstract: Existing benchmarks typically report accuracy for a single model on a single run. This systematically understates real-world LLM capabilities, particula

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

The Geometry of Updates: Fisher Alignment at Vocabulary Scale

DGX agent

arXiv:2606.27242v1 Announce Type: cross Abstract: Training-free source selection for LLM families with shared vocabularies arises in scientific string domains such as SMILES, protein, and genomic sequ

model-releasesarxiv-cs-cl
26 Jun 2026
Model Releases

The Inattentional Gap: Task-Conditioned Language and Vision Models Omit the Safety-Critical Signals They Can Otherwise Report

DGX agent

arXiv:2606.26529v1 Announce Type: cross Abstract: AI safety is evaluated by how reliably a model detects the hazards it is told to find, yet accidents often arise from the hazard no one specified. We

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

The Open Source Economic Index of AI Adoption and Capability

DGX agent

arXiv:2606.26118v1 Announce Type: cross Abstract: We work towards measuring both AI adoption and the capability of AI to perform discrete labor tasks across various occupations. To measure adoption, w

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

The Red Queen Godel Machine: Co-Evolving Agents and Their Evaluators

DGX agent

arXiv:2606.26294v1 Announce Type: cross Abstract: Self-improving agents are state-of-the-art (SOTA) on agentic coding benchmarks and have recently been extended to general domains. However, their sear

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

The Spec Growth Engine: Spec-Anchored, Code-Coupled, Drift-Enforced Architecture for AI-Assisted Software Development

DGX agent

arXiv:2606.27045v1 Announce Type: cross Abstract: AI coding agents dramatically accelerate implementation speed but introduce two structural failure modes that existing spec-driven approaches do not f

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

The strongest models are gated and access is granted only to a select few. Hermes Agent now exposes MoA presets as virtual models, giving yo…

DGX agent

The strongest models are gated and access is granted only to a select few. Hermes Agent now exposes MoA presets as virtual models, giving you capabilities beyond the publicly available frontier: 8% hi

model-releasesnous-research--x
26 Jun 2026
Model Releases

Theory-Scale Auto-Formalization of Logics for Computer Science

DGX agent

arXiv:2606.26525v1 Announce Type: new Abstract: Auto-formalization is critical for scalable formal verification, but existing progress largely focuses on isolated statements, while theory-scale auto-f

model-releasesarxiv-cs-lg
26 Jun 2026
Model Releases

Thinking Like a Scientist? A Structural Study of LLM-Generated Research Methods

DGX agent

arXiv:2606.26130v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used to guide research methodology, yet their default methodological tendencies under minimal prompting

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

This week someone targeted my family for harm with a false report. We’re physically OK, but that doesn’t mean we weren’t harmed. I am beyond…

DGX agent

This week someone targeted my family for harm with a false report. We’re physically OK, but that doesn’t mean we weren’t harmed. I am beyond furious. Whatever your politics, this is awful, wrong, and

model-releasesanthropic--x
26 Jun 2026
Model Releases

TMP: Tree-structured Mixed-policy Pruning for Large-scale Image Generation and Editing

DGX agent

arXiv:2606.27089v1 Announce Type: new Abstract: Modern image generation model rapidly grows their sizes to meet high-fidelity image synthesis. However, they gradually become unaffordable for their eno

model-releasesarxiv-cs-cv
26 Jun 2026
Model Releases

Today on TITV: -Trump asks OpenAI to stagger release of new model | @leomschwartz & @amir, The Information -Google pressures publishers on A…

DGX agent

Today on TITV: -Trump asks OpenAI to stagger release of new model | @leomschwartz & @amir, The Information -Google pressures publishers on AI licensing | @anngehan -Inside an AI power user’s agent wor

model-releasesallie-k--miller--x
26 Jun 2026
Model Releases

Towards Explainable Adjudicative Variance: Quantifying Judicial Discretion via Gated Multi-Task Learning

DGX agent

arXiv:2606.27069v1 Announce Type: new Abstract: Legal outcome prediction must disentangle objective case facts from adjudicative context. Merit-based rulings rely on factual evidence while technical d

model-releasesarxiv-cs-cl
26 Jun 2026
← Previous
1…162163164165166…471
Next →