AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,585 results
Model Releases

WATCH: Wide-Area Archaeological Site Tracking for Change Detection

DGX agent

arXiv:2605.08160v1 Announce Type: cross Abstract: Monitoring archaeological sites at scale is vital for protecting cultural heritage, yet pinpointing when disturbances occur remains difficult because

model-releasesarxiv-cs-ai
12 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

We integrated FrontierCS into Harbor and are releasing a preview long-horizon agent leaderboard (up to 835 turns, ~200K output tokens) with …

DGX agent

We integrated FrontierCS into Harbor and are releasing a preview long-horizon agent leaderboard (up to 835 turns, ~200K output tokens) with Kimi K2.6 @Kimi_Moonshot (score 46.9) and Claude Code Opus 4

model-releaseskimi-moonshot--x
12 May 2026
Model Releases

Weight Pruning Amplifies Bias: A Multi-Method Study of Compressed LLMs for Edge AI

DGX agent

arXiv:2605.08137v1 Announce Type: cross Abstract: Weight pruning is widely advocated for deploying Large Language Models on resource-constrained IoT and edge devices, yet its impact on model fairness

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

What Parameter Golf taught us about AI-assisted research

DGX agent

Parameter Golf brought together 1,000+ participants and 2,000+ submissions to explore AI-assisted machine learning research, coding agents, quantization, and novel model design under strict constraint

model-releasesopenai
12 May 2026
Model Releases

What Will Happen Next: Large Models-Driven Deduction for Emergency Instances

DGX agent

arXiv:2605.08599v1 Announce Type: new Abstract: Traditional simulation methods reproduce occurred emergency instances through presetting to assist people in risk assessment and emergency decision-maki

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

What’s new in Microsoft Foundry | April 2026

DGX agent

April brings Foundry Local GA for local AI development, GPT-5.5 model support with Tier 5 and Tier 6 default quota in Microsoft Foundry, new tracing paths for Microsoft Agent Framework and hosted agen

model-releasesmicrosoft-foundry
12 May 2026
Model Releases

What's the plan? Metrics for implicit planning in LLMs and their application to rhyme generation and question answering

DGX agent

arXiv:2601.20164v2 Announce Type: replace-cross Abstract: Prior work suggests that language models, while trained on next token prediction, show implicit planning behavior: they may select the next to

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

When Adaptation Fails: A Gradient-Based Diagnosis of Collapsed Gating in Vision-Language Prompt Learning

DGX agent

arXiv:2605.09549v1 Announce Type: new Abstract: Adaptive prompting mechanisms have been proposed to enhance vision-language models by dynamically tailoring prompts to inputs. However, in frozen few-sh

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

When (and How) to Trust the Expert: Diagnosing Query-Time Expert-Guided Reinforcement Learning

DGX agent

arXiv:2605.09109v1 Announce Type: new Abstract: Many continuous-control problems ship with a competent but suboptimal controller (a tuned PID, a hand-designed gait). A growing family of methods uses s

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

When Attention Beats Fourier: Multi-Scale Transformers for PDE Solving on Irregular Domains

DGX agent

arXiv:2605.08318v1 Announce Type: cross Abstract: We study the problem of architecture selection for deep learning models trained to solve partial differential equations (PDEs), asking when transforme

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

When Does Non-Uniform Replay Matter in Reinforcement Learning?

DGX agent

arXiv:2605.10236v1 Announce Type: cross Abstract: Modern off-policy reinforcement learning algorithms often rely on simple uniform replay sampling and it remains unclear when and why non-uniform repla

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

When is the last time a general purpose LLM (putting aside hybrid systems like Claude Code with special purpose symbolic harnesses) last com…

DGX agent

When is the last time a general purpose LLM (putting aside hybrid systems like Claude Code with special purpose symbolic harnesses) last completely blew away all competing prior models? GPT-4 relative

model-releasesgary-marcus--x
12 May 2026
Model Releases

When Prompts Become Payloads: A Framework for Mitigating SQL Injection Attacks in Large Language Model-Driven Applications

DGX agent

arXiv:2605.10176v1 Announce Type: cross Abstract: Natural language interfaces to structured databases are becoming increasingly common, largely due to advances in large language models (LLMs) that ena

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

When Reviews Disagree: Fine-Grained Contradiction Analysis in Scientific Peer Reviews

DGX agent

arXiv:2605.10171v1 Announce Type: cross Abstract: Scientific peer reviews frequently contain conflicting expert judgments, and the increasing scale of conference submissions makes it challenging for A

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

When to Re-Commit: Temporal Abstraction Discovery for Long-Horizon Vision-Language Reasoning

DGX agent

arXiv:2605.09860v1 Announce Type: new Abstract: Long-horizon reasoning requires deciding not only what actions to take, but how deeply to commit before the next observation. We formalize this as commi

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

When to Trust Imagination: Adaptive Action Execution for World Action Models

DGX agent

arXiv:2605.06222v2 Announce Type: replace-cross Abstract: World Action Models (WAMs) have recently emerged as a promising paradigm for robotic manipulation by jointly predicting future visual observat

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Where Does Long-Context Supervision Actually Go? Effective-Context Exposure Balancing

DGX agent

arXiv:2605.10544v1 Announce Type: new Abstract: Long-context adaptation is often viewed as window scaling, but this misses a token-level supervision mismatch: in packed training with document masking,

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

Why Retrying Fails: Context Contamination in LLM Agent Pipelines

DGX agent

arXiv:2605.08563v1 Announce Type: new Abstract: When an LLM agent fails a multi-step tool-augmented task and retries, the failed attempt typically remains in its context window -- contaminating the ne

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Why Zeroth-Order Adaptation May Forget Less: A Randomized Shaping Theory

DGX agent

arXiv:2605.10658v1 Announce Type: new Abstract: Continual learning requires new-task adaptation without damaging previously acquired capabilities. Recent forward-pass and zeroth-order (ZO) results sho

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

WildClawBench: A Benchmark for Real-World, Long-Horizon Agent Evaluation

DGX agent

arXiv:2605.10912v1 Announce Type: new Abstract: Large language and vision-language models increasingly power agents that act on a user's behalf through command-line interface (CLI) harnesses. However,

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

WindINR: Latent-State INR for Fast Local Wind Query and Correction in Complex Terrain

DGX agent

arXiv:2605.09511v1 Announce Type: new Abstract: Many downstream decisions in complex terrain require fast wind estimates at a small number of user-specified locations and heights for a given forecast

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Wordle 1,787 3/6 ⬛⬛⬛⬛🟨 🟨⬛🟨⬛⬛ 🟩🟩🟩🟩🟩

DGX agent

This post documents a Wordle game result where the player solved puzzle #1,787 in three attempts, using the color-coded emoji system to show letter placement feedback (black for wrong letters, yellow

model-releasesanthropic--x
12 May 2026
Model Releases

Wordle 1,788 5/6 ⬛⬛⬛⬛⬛ 🟨⬛⬛⬛🟨 ⬛🟨🟨⬛⬛ 🟩🟩🟩⬛⬛ 🟩🟩🟩🟩🟩

DGX agent

This post documents a completed game of Wordle (puzzle #1,788) played by Anthropic, showing the progression of guesses across six attempts with color-coded feedback indicating correct letter positions

model-releasesanthropic--x
12 May 2026
Model Releases

WorldReasonBench: Human-Aligned Stress Testing of Video Generators as Future World-State Predictors

DGX agent

arXiv:2605.10434v1 Announce Type: new Abstract: Commercial video generation systems such as Seedance2.0 and Veo3.1 have rapidly improved, strengthening the view that video generators may be evolving i

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

You Have Been LaTeXpOsEd: A Systematic Analysis of Information Leakage in Preprint Archives Using Large Language Models

DGX agent

arXiv:2510.03761v2 Announce Type: replace-cross Abstract: The widespread use of preprint repositories such as arXiv has accelerated the communication of scientific results but also introduced overlook

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Your Simulation Runs but Solves the Wrong Physics: PDE-Grounded Intent Verification for LLM-Generated Multiphysics Simulation Code

DGX agent

arXiv:2605.09360v1 Announce Type: cross Abstract: Execution-based evaluation of LLM-generated code implicitly treats successful execution as a proxy for correctness. In scientific simulation, this pro

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Zero-Shot Chinese Character Recognition via Global-Local Dual-Branch Alignment and Hierarchical Inference

DGX agent

arXiv:2605.08814v1 Announce Type: new Abstract: Chinese character categories are extremely large, and unseen characters frequently arise in open-world scenarios, making zero-shot Chinese character rec

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

+1 to this. I was recently on a cross-continental flight without wifi, so I brought up Qwen3.6 & Gemma 4 (via @ollama) in Deep Agents on my …

DGX agent

+1 to this. I was recently on a cross-continental flight without wifi, so I brought up Qwen3.6 & Gemma 4 (via @ollama) in Deep Agents on my laptop. admittedly, they fell over on some more involved/com

model-releasesharrison-chase--x
11 May 2026
Model Releases

2.5-D Decomposition for LLM-Based Spatial Construction

DGX agent

arXiv:2605.07066v1 Announce Type: new Abstract: Autonomous systems that build structures from natural-language instructions need reliable spatial reasoning, yet large language models (LLMs) make syste

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

3 weeks since ml-intern launched and we just hit 1M messages exchanged. that's 3.3 agent-years of ML research in 21 days. 2 months worth of …

DGX agent

3 weeks since ml-intern launched and we just hit 1M messages exchanged. that's 3.3 agent-years of ML research in 21 days. 2 months worth of research every day. 17,383 training jobs total. talk about A

model-releasesclem-delangue--x
11 May 2026
Model Releases

A Causal Diffusion Model for Video Reconstruction from Ultra-Low-Bitrate Representations

DGX agent

arXiv:2602.13837v2 Announce Type: replace Abstract: We study video reconstruction from ultra-low-bitrate representations, where the primary challenge shifts from encoding to decoding. In this regime,

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

A Hierarchical Ensemble Pipeline for Anomaly Detection in ESA Satellite Telemetry

DGX agent

arXiv:2605.06681v1 Announce Type: cross Abstract: A hierarchical ensemble pipeline is introduced to address anomaly detection in multivariate telemetry data provided by European Space Agency (ESA). Th

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

A Marine Debris Detection Framework for Ocean Robots via Self-Attention Enhancement and Feature Interaction Optimization

DGX agent

arXiv:2605.07388v1 Announce Type: new Abstract: Marine debris detection for ocean robot is crucial for ecological protection, yet performance is often degraded by low-quality images with blur, complex

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

A Reproducible Multi-Architecture Baseline for Token-Level Chinese Metaphor Identification under the MIPVU Framework

DGX agent

arXiv:2605.07170v1 Announce Type: new Abstract: Metaphor is pervasive in everyday language, yet token-level computational identification of metaphor-related words in Chinese under the MIPVU framework

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

A Reproducible Optimisation Protocol for Calibrating Prompt-Based Large Language Model Workflows in Evidence Synthesis

DGX agent

arXiv:2605.06937v1 Announce Type: new Abstract: This methods article presents a reproducible calibration workflow for prompt-based large language models (LLMs) in structured evidence-synthesis tasks.

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

A Unified and Controllable Framework for Layered Image Generation with Visual Effects

DGX agent

arXiv:2601.15507v2 Announce Type: replace Abstract: Recent image generation models produce impressive composites, but often fail to preserve the identity of user-provided content when editing specific

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

A^2RD: Agentic Autoregressive Diffusion for Long Video Consistency

DGX agent

arXiv:2605.06924v1 Announce Type: cross Abstract: Synthesizing consistent and coherent long video remains a fundamental challenge. Existing methods suffer from semantic drift and narrative collapse ov

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Adapting Vision-Language Models for Neutrino Event Classification in High-Energy Physics

DGX agent

arXiv:2509.08461v3 Announce Type: replace-cross Abstract: Recent advances in Large Language Models (LLMs) have demonstrated their remarkable capacity to process and reason over structured and unstruct

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Adaptive Memory Decay for Log-Linear Attention

DGX agent

arXiv:2605.06946v1 Announce Type: cross Abstract: Sequence models face a fundamental tradeoff between memory capacity and computational efficiency. Transformers achieve expressive context modeling at

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Adaptive Regularization for Sparsity Control in Bregman-Based Optimizers

DGX agent

arXiv:2605.07892v1 Announce Type: new Abstract: Sparse training reduces the memory and computational costs of deep neural networks. However, sparse optimization methods, e.g., those adding an ell_1 pe

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

Agent view is the best Claude Code native way to manage multiple sessions, kind of like tmux built for CC. We spent a lot of time getting th…

DGX agent

Agent view is the best Claude Code native way to manage multiple sessions, kind of like tmux built for CC. We spent a lot of time getting the details right, I hope you enjoy it. New in Claude Code: ag

model-releasesthariq--x
11 May 2026
Model Releases

AgentEscapeBench: Evaluating Out-of-Domain Tool-Grounded Reasoning in LLM Agents

DGX agent

arXiv:2605.07926v1 Announce Type: new Abstract: As LLM-based agents increasingly rely on external tools, it is important to evaluate their ability to sustain tool-grounded reasoning beyond familiar wo

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Agentick: A Unified Benchmark for General Sequential Decision-Making Agents

DGX agent

arXiv:2605.06869v1 Announce Type: new Abstract: AI agent research spans a wide spectrum: from RL agents that learn from scratch to foundation model agents that leverage pre-trained knowledge, yet no u

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Agents + file sandboxes are all in the range in 2026 🤖🗃️ This is a nifty reference implementation by @itsclelia showing you how to run you…

DGX agent

Agents + file sandboxes are all in the range in 2026 🤖🗃️ This is a nifty reference implementation by @itsclelia showing you how to run your agent over a collection of docs (PDFs, images, Office) with

model-releasesjerry-liu--x
11 May 2026
Model Releases

AI CFD Scientist: Toward Open-Ended Computational Fluid Dynamics Discovery with Physics-Aware AI Agents

DGX agent

arXiv:2605.06607v2 Announce Type: replace-cross Abstract: Recent LLM-based agents have closed substantial portions of the scientific discovery loop in software-only machine-learning research, in chemi

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

@Alibaba_Qwen Sign up to get free access to Qwen 3.6 Plus and much more at http://portal.nousresearch.com/manage-subscription!

DGX agent

Nous Research announced free access to Qwen 3.6 Plus and additional features available through their subscription portal at portal.nousresearch.com. This promotion likely provides users with complimen

model-releasesnous-research--x
11 May 2026
Model Releases

Amortized Molecular Optimization via Group Relative Policy Optimization

DGX agent

arXiv:2602.12162v3 Announce Type: replace Abstract: In structurally constrained molecular optimization, state-of-the-art methods restart an expensive oracle-driven search from scratch for every new in

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

Amortized Multi-Objective Optimization Across Tasks with Generative Solution Modeling

DGX agent

arXiv:2511.09598v5 Announce Type: replace Abstract: Many real-world applications require solving families of expensive multi-objective optimization problems~(EMOPs) under varying operational condition

model-releasesarxiv-cs-lg
11 May 2026
← Previous
1…338339340341342…471
Next →