AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,115
  • Agents7,313
  • Applications5,228
  • Concepts5
  • Hardware1,762
  • Industry6,105
  • Local Ai4,756
  • Model Releases22,759
  • Research19,333
  • Safety12,889
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,115
  • Agents7,313
  • Applications5,228
  • Concepts5
  • Hardware1,762
  • Industry6,105
  • Local Ai4,756
  • Model Releases22,759
  • Research19,333
  • Safety12,889
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
85,115Total entries
1Added by human
85,114Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,767 results
Model Releases

SharQ: Bridging Activation Sparsity and FP4 Quantization for LLM Inference

DGX agent

arXiv:2606.26587v1 Announce Type: cross Abstract: Low-bit floating-point formats and semi-structured sparsity are increasingly supported by modern accelerators, yet combining them for LLM activation c

model-releasesarxiv-cs-ai
26 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Signature filtering: a lightweight enhancement for statistical watermark detection in large language models

DGX agent

arXiv:2606.18430v2 Announce Type: replace Abstract: Statistical watermarks help organizations attribute large language model (LLM) outputs, yet existing detectors often struggle when watermark signals

model-releasesarxiv-cs-lg
26 Jun 2026
Model Releases

Simulation-based inference for rapid Bayesian parameter estimation in epidemiological models: a comparison with MCMC

DGX agent

arXiv:2606.27286v1 Announce Type: new Abstract: Mechanistic epidemiological models are widely used to support infectious disease forecasting and public-health decision making. Bayesian calibration of

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

SITUATION EXPLAINED: Anthropic accused Alibaba of conducting a 'distillation attack' on Claude. We asked @ClementDelangue his thoughts: 'Dis…

DGX agent

SITUATION EXPLAINED: Anthropic accused Alibaba of conducting a 'distillation attack' on Claude. We asked @ClementDelangue his thoughts: 'Distillation is a very common practice that everyone is using.

model-releasesclem-delangue--x
26 Jun 2026
Model Releases

SITUATION EXPLAINED: What is at stake in the open source AI race? We asked @perrymetzger, chairman of Alliance for the Future. 'Open source …

DGX agent

SITUATION EXPLAINED: What is at stake in the open source AI race? We asked @perrymetzger, chairman of Alliance for the Future. 'Open source models are a sort of cultural ambassador and civilizational

model-releasesclem-delangue--x
26 Jun 2026
Model Releases

SocialPersona: Benchmarking Personalized Profiling and Response with Multimodal Social-Media Context

DGX agent

arXiv:2606.26654v1 Announce Type: new Abstract: Personalized language-model assistants are often evaluated through a memory lens: can a model recall preferences users have explicitly stated in dialogu

model-releasesarxiv-cs-cl
26 Jun 2026
Model Releases

Sources: DeepSeek's $7.4B raise was prompted by the release of Mythos as CEO Liang Wenfeng realized DeepSeek couldn't compete without a massive war chest (The Information)

DGX agent

The Information: Sources: DeepSeek's $7.4B raise was prompted by the release of Mythos as CEO Liang Wenfeng realized DeepSeek couldn't compete without a massive war chest — Up until two months ago, De

model-releasestechmeme
26 Jun 2026
Model Releases

Spurious Rewards Paradox: Mechanistically Understanding How RLVR Activates Memorization Shortcuts in LLMs

DGX agent

arXiv:2601.11061v2 Announce Type: replace-cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) is highly effective for enhancing LLM reasoning, yet recent evidence shows models like Q

model-releasesarxiv-cs-cl
26 Jun 2026
Model Releases

SSI-Policy: Learning Structured Scene Interfaces for Vision-Language Robotic Manipulation

DGX agent

arXiv:2606.26800v1 Announce Type: new Abstract: Real-world robotic manipulation demands spatial grounding, task-aware reasoning, and precise control. Learning such capabilities becomes particularly ch

model-releasesarxiv-cs-ro
26 Jun 2026
Model Releases

SSM Adapters via Hankel Reduced-order Modeling: Injection Site Determines Task Suitability in Long-Context Fine-Tuning

DGX agent

arXiv:2606.26290v1 Announce Type: cross Abstract: While parameter-efficient fine-tuning (PEFT) typically targets attention projectors, its efficacy for tasks requiring sequential state accumulation re

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

Stochastic Gradient Optimization with Model-Assisted Sampling

DGX agent

arXiv:2606.27171v1 Announce Type: new Abstract: This work addresses the problem of variance in stochastic gradient estimation for machine learning optimization. Deep learning relies on mini-batch meth

model-releasesarxiv-cs-lg
26 Jun 2026
Model Releases

Taps sign

DGX agent

Taps sign We believe in broad access and plan to make GPT-5.6 Sol, Terra, and Luna generally available in the coming weeks. For now, at the request of the U.S. government, we’re starting with a limite

model-releasescohere--x
26 Jun 2026
Model Releases

Temporally Consistent Label Interpolation for Robust Surgical Multi-Task Learning under Challenging Conditions

DGX agent

arXiv:2606.26634v1 Announce Type: new Abstract: Effective multi-task learning for surgical scene understanding is fundamentally hindered by annotation granularity mismatch; temporal workflow tasks suc

model-releasesarxiv-cs-cv
26 Jun 2026
Model Releases

Term-Centric Hierarchy Induction from Heterogeneous Corpora

DGX agent

arXiv:2606.26963v1 Announce Type: new Abstract: Organizing knowledge from diverse text sources into interpretable hierarchies is crucial for tasks such as policy analysis, innovation monitoring, and e

model-releasesarxiv-cs-cl
26 Jun 2026
Model Releases

TF-TI2I: Training-Free Text-and-Image-to-Image Generation via Multi-Modal Implicit-Context Learning in Text-to-Image Models

DGX agent

arXiv:2503.15283v2 Announce Type: replace Abstract: Text-and-Image-To-Image (TI2I), an extension of Text-To-Image (T2I), integrates image inputs with textual instructions to enhance image generation.

model-releasesarxiv-cs-cv
26 Jun 2026
Model Releases

The Capability Frontier: Benchmarks Miss 82% of Model Performance

DGX agent

arXiv:2606.26836v1 Announce Type: new Abstract: Existing benchmarks typically report accuracy for a single model on a single run. This systematically understates real-world LLM capabilities, particula

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

The Geometry of Updates: Fisher Alignment at Vocabulary Scale

DGX agent

arXiv:2606.27242v1 Announce Type: cross Abstract: Training-free source selection for LLM families with shared vocabularies arises in scientific string domains such as SMILES, protein, and genomic sequ

model-releasesarxiv-cs-cl
26 Jun 2026
Model Releases

The Inattentional Gap: Task-Conditioned Language and Vision Models Omit the Safety-Critical Signals They Can Otherwise Report

DGX agent

arXiv:2606.26529v1 Announce Type: cross Abstract: AI safety is evaluated by how reliably a model detects the hazards it is told to find, yet accidents often arise from the hazard no one specified. We

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

The Open Source Economic Index of AI Adoption and Capability

DGX agent

arXiv:2606.26118v1 Announce Type: cross Abstract: We work towards measuring both AI adoption and the capability of AI to perform discrete labor tasks across various occupations. To measure adoption, w

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

The Red Queen Godel Machine: Co-Evolving Agents and Their Evaluators

DGX agent

arXiv:2606.26294v1 Announce Type: cross Abstract: Self-improving agents are state-of-the-art (SOTA) on agentic coding benchmarks and have recently been extended to general domains. However, their sear

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

The Spec Growth Engine: Spec-Anchored, Code-Coupled, Drift-Enforced Architecture for AI-Assisted Software Development

DGX agent

arXiv:2606.27045v1 Announce Type: cross Abstract: AI coding agents dramatically accelerate implementation speed but introduce two structural failure modes that existing spec-driven approaches do not f

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

The strongest models are gated and access is granted only to a select few. Hermes Agent now exposes MoA presets as virtual models, giving yo…

DGX agent

The strongest models are gated and access is granted only to a select few. Hermes Agent now exposes MoA presets as virtual models, giving you capabilities beyond the publicly available frontier: 8% hi

model-releasesnous-research--x
26 Jun 2026
Model Releases

Theory-Scale Auto-Formalization of Logics for Computer Science

DGX agent

arXiv:2606.26525v1 Announce Type: new Abstract: Auto-formalization is critical for scalable formal verification, but existing progress largely focuses on isolated statements, while theory-scale auto-f

model-releasesarxiv-cs-lg
26 Jun 2026
Model Releases

Thinking Like a Scientist? A Structural Study of LLM-Generated Research Methods

DGX agent

arXiv:2606.26130v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used to guide research methodology, yet their default methodological tendencies under minimal prompting

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

This week someone targeted my family for harm with a false report. We’re physically OK, but that doesn’t mean we weren’t harmed. I am beyond…

DGX agent

This week someone targeted my family for harm with a false report. We’re physically OK, but that doesn’t mean we weren’t harmed. I am beyond furious. Whatever your politics, this is awful, wrong, and

model-releasesanthropic--x
26 Jun 2026
Model Releases

TMP: Tree-structured Mixed-policy Pruning for Large-scale Image Generation and Editing

DGX agent

arXiv:2606.27089v1 Announce Type: new Abstract: Modern image generation model rapidly grows their sizes to meet high-fidelity image synthesis. However, they gradually become unaffordable for their eno

model-releasesarxiv-cs-cv
26 Jun 2026
Model Releases

Today on TITV: -Trump asks OpenAI to stagger release of new model | @leomschwartz & @amir, The Information -Google pressures publishers on A…

DGX agent

Today on TITV: -Trump asks OpenAI to stagger release of new model | @leomschwartz & @amir, The Information -Google pressures publishers on AI licensing | @anngehan -Inside an AI power user’s agent wor

model-releasesallie-k--miller--x
26 Jun 2026
Model Releases

Towards Explainable Adjudicative Variance: Quantifying Judicial Discretion via Gated Multi-Task Learning

DGX agent

arXiv:2606.27069v1 Announce Type: new Abstract: Legal outcome prediction must disentangle objective case facts from adjudicative context. Merit-based rulings rely on factual evidence while technical d

model-releasesarxiv-cs-cl
26 Jun 2026
Model Releases

Towards Video Anomaly Detection from Event Streams: A Baseline and Benchmark Datasets

DGX agent

arXiv:2603.24991v2 Announce Type: replace Abstract: Event-based vision, characterized by low redundancy, focus on dynamic motion, and inherent privacy-preserving properties, naturally fits the demands

model-releasesarxiv-cs-cv
26 Jun 2026
Model Releases

TraMP-LLaMA: Generative Interpretability with Decoupled Instruction Tuning for Facial Expression Quality Assessment

DGX agent

arXiv:2606.26942v1 Announce Type: new Abstract: Existing facial expression quality assessment (FEQA) methods typically produce only a severity score, without explicitly communicating the observable fa

model-releasesarxiv-cs-cv
26 Jun 2026
Model Releases

Tuning Language Models by Mixture-of-Depths Ensemble

DGX agent

arXiv:2410.13077v2 Announce Type: replace-cross Abstract: Transformer-based Large Language Models (LLMs) traditionally rely on final-layer loss for finetuning and final-layer representations for predi

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

UAV-MapFusion: RTK-Aligned Uncertainty-Aware Coarse-to-Fine Multi-Session UAV Mapping

DGX agent

arXiv:2606.26928v1 Announce Type: new Abstract: Large-scale point cloud maps are essential for robotics and spatial intelligence tasks. UAVs provide an efficient means for large-scale map acquisition;

model-releasesarxiv-cs-ro
26 Jun 2026
Model Releases

UBS says 60% of companies now watching AI budgets are moving to cheaper models and open-source Chinese models The pressure is coming from ex…

DGX agent

UBS says 60% of companies now watching AI budgets are moving to cheaper models and open-source Chinese models The pressure is coming from extreme bills, including users spending up to $35K/month, team

model-releasesclem-delangue--x
26 Jun 2026
Model Releases

Unconventional AI debuts oscillator-based Un-0 model series

DGX agent

Unconventional AI Inc. has developed an artificial intelligence architecture that could improve the power efficiency of image generation models. The technology is the basis of a new neural network ser

model-releasessiliconangle
26 Jun 2026
Model Releases

Unison: Benchmarking Unified Multimodal Models via Synergistic Understanding and Generation

DGX agent

arXiv:2606.26984v1 Announce Type: new Abstract: Unified multimodal models capable of both understanding and generation have achieved remarkable strides. However, despite their unified designs, existin

model-releasesarxiv-cs-cv
26 Jun 2026
Model Releases

VisNec: Measuring and Leveraging Visual Necessity for Multimodal Instruction Tuning

DGX agent

arXiv:2603.01195v2 Announce Type: replace-cross Abstract: The effectiveness of multimodal instruction tuning depends not only on dataset scale, but critically on whether training samples genuinely req

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

Visual-Language-Guided Task Planning for Horticultural Robots

DGX agent

arXiv:2601.11906v2 Announce Type: replace Abstract: Crop monitoring is essential for precision agriculture, but current systems lack high-level reasoning. We introduce a novel, modular framework that

model-releasesarxiv-cs-ro
26 Jun 2026
Model Releases

Vulnerability of Natural Language Classifiers to Evolutionary Generated Adversarial Text

DGX agent

arXiv:2606.27215v1 Announce Type: new Abstract: Deep learning models have achieved impressive performance across various fields but remain vulnerable to adversarial inputs, particularly in NLP, where

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

WatchAct: A Benchmark for Behavior-Grounded Robot Manipulation

DGX agent

arXiv:2606.26443v1 Announce Type: cross Abstract: A robot working alongside people must reason about what they have done, in what order, and with what intent. Video carries the spatial layouts, object

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

We stress tested many frontier AI models for multimodal medical reasoning (including GPT-5, Claude 3.5, Gemini 2.5 Pro). They’re not ready. …

DGX agent

We stress tested many frontier AI models for multimodal medical reasoning (including GPT-5, Claude 3.5, Gemini 2.5 Pro). They’re not ready. Faulty reasoning, use of inappropriate shortcuts, hallucinat

model-releasesgary-marcus--x
26 Jun 2026
Model Releases

What Do Deepfake Benchmarks Measure? An Audit Using Frozen Self-Supervised Representations

DGX agent

arXiv:2606.26384v1 Announce Type: new Abstract: As deepfake generators approach perceptual indistinguishability, reliable detection becomes critical. Yet, detectors that score well on benchmarks routi

model-releasesarxiv-cs-cv
26 Jun 2026
Model Releases

What happened after 2,000 people tried to hack my AI assistant

DGX agent

What happened after 2,000 people tried to hack my AI assistant Fernando Irarrázaval ran a challenge on hackmyclaw.com to see if anyone could leak secrets held by his OpenClaw test instance by sending

model-releasessimon-willison
26 Jun 2026
Model Releases

What We are Missing in Multimodal LLM Evaluation?

DGX agent

arXiv:2606.26348v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) can process diverse inputs, e.g., text, images, audio, and video, and generate textual responses. While their c

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

When Actions Go Off-Task: Detecting and Correcting Misaligned Actions in Computer-Use Agents

DGX agent

arXiv:2602.08995v2 Announce Type: replace Abstract: Computer-use agents (CUAs) have made tremendous progress in the past year, yet they still frequently produce misaligned actions that deviate from th

model-releasesarxiv-cs-cl
26 Jun 2026
Model Releases

When to Write and When to Suppress: Route-Specialized Dual Adapters for Memory-Assisted Knowledge Editing

DGX agent

arXiv:2606.14668v3 Announce Type: replace Abstract: Knowledge editing systems must update selected facts while preserving nearby but irrelevant behavior. This paper studies this problem in a memory-as

model-releasesarxiv-cs-lg
26 Jun 2026
Model Releases

Where Do Models Find Happiness? Emotion Vectors in Open-Source LLMs

DGX agent

arXiv:2606.26987v1 Announce Type: cross Abstract: Recent work identified emotion vectors in Claude Sonnet 4.5, which are internal representations that encode emotion concepts, causally influence behav

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

Which one is Claude is pretty obvious. GLM-5.2 is a beast in some ways, but doesn't have the self-reflective persona of Claude, and isn't re…

DGX agent

Ethan Mollick compares Claude and GLM-5.2 AI models, noting that while GLM-5.2 excels in certain capabilities, Claude distinguishes itself through its self-reflective persona and other characteristics

model-releasesethan-mollick--x
26 Jun 2026
Model Releases

Wordle 1,832 4/6 ⬛🟨⬛⬛⬛ ⬛🟨🟨⬛⬛ 🟨⬛🟩🟨⬛ 🟩🟩🟩🟩🟩

DGX agent

This post documents a Wordle game result where the player solved puzzle #1,832 in four attempts, using the color-coded emoji system (gray for incorrect letters, yellow for correct letters in wrong pos

model-releasesanthropic--x
26 Jun 2026
← Previous
1…167168169170171…475
Next →