AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,569 results
Model Releases

When Prompts Interact: Assessing Prompt Arithmetic for Deconfounding under Distribution Shift

DGX agent

arXiv:2605.03096v1 Announce Type: cross Abstract: In classification tasks, models may rely on confounding variables to achieve strong in-distribution performance, capturing spurious features that fail

model-releasesarxiv-cs-cl
6 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

When Should a Language Model Trust Itself? Same-Model Self-Verification as a Conditional Confidence Signal

DGX agent

arXiv:2605.02915v1 Announce Type: new Abstract: Same-model self-verification, prompting a model to audit its own predicted answer, is a plausible confidence signal for selective prediction, but its pr

model-releasesarxiv-cs-cl
6 May 2026
Model Releases

When Stress Becomes Signal: Detecting Antifragility-Compatible Regimes in Multi-Agent LLM Systems

DGX agent

arXiv:2605.02463v1 Announce Type: cross Abstract: Multi-agent LLM systems are increasingly used to solve complex tasks through decomposition, debate, specialization, and ensemble reasoning. However, t

model-releasesarxiv-cs-ai
6 May 2026
Model Releases

Where to Bind Matters: Hebbian Fast Weights in Vision Transformers for Few-Shot Character Recognition

DGX agent

arXiv:2605.02920v1 Announce Type: cross Abstract: Standard transformer architectures learn fixed slow-weight representations during training and lack mechanisms for rapid adaptation within an episode.

model-releasesarxiv-cs-cv
6 May 2026
Model Releases

Wordle 1,781 3/6 ⬛🟨🟨⬛⬛ 🟩⬛⬛⬛⬛ 🟩🟩🟩🟩🟩

DGX agent

Anthropic shared their Wordle game result for puzzle #1,781, solved in 3 attempts with a final correct answer shown by five green squares (all letters in correct positions). The emoji grid represents

model-releasesanthropic--x
6 May 2026
Model Releases

Workspace-Bench 1.0: Benchmarking AI Agents on Workspace Tasks with Large-Scale File Dependencies

DGX agent

arXiv:2605.03596v1 Announce Type: cross Abstract: Workspace learning requires AI agents to identify, reason over, exploit, and update explicit and implicit dependencies among heterogeneous files in a

model-releasesarxiv-cs-cl
6 May 2026
Model Releases

WorldJen: An End-to-End Multi-Dimensional Benchmark for Generative Video Models

DGX agent

arXiv:2605.03475v1 Announce Type: new Abstract: Evaluating generative video models remains an open problem. Reference-based metrics such as Structural Similarity Index Measure (SSIM) and Peak Signal t

model-releasesarxiv-cs-cv
6 May 2026
Model Releases

xAI and SpaceXAI have just made Colossus 1 available to Anthropic to support Claude. This means more than 220,000 NVIDIA GPUs in one of the …

DGX agent

xAI and SpaceXAI have just made Colossus 1 available to Anthropic to support Claude. This means more than 220,000 NVIDIA GPUs in one of the world’s largest and fastest-built AI superclusters are now h

model-releaseselon-musk--x
6 May 2026
Model Releases

A Closed-Form Persistence-Landmark Pipeline for Certified Point-Cloud and Graph Classification

DGX agent

arXiv:2605.02836v1 Announce Type: new Abstract: We introduce PLACE (Persistence-Landmark Analytic Classification Engine), a closed-form pipeline for classifying point clouds and graphs through their p

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

A decoupled diffusion planner that adapts to changing cost limits by using cost-conditioned generation for safety and reward gradients for performance

DGX agent

arXiv:2605.02777v1 Announce Type: new Abstract: Offline safe reinforcement learning often requires policies to adapt at deployment time to safety budgets that vary across episodes or change within a s

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

A few days ago i wrote a post about how we use the LangGraph checkpointer to optimize storage. Today, LangGraph released v1.2 with a really …

DGX agent

A few days ago i wrote a post about how we use the LangGraph checkpointer to optimize storage. Today, LangGraph released v1.2 with a really nice feature: DeltaChannel, a new channel type that stores o

model-releasesharrison-chase--x
5 May 2026
Model Releases

A hybrid solution approach for the Integrated Healthcare Timetabling Competition 2024

DGX agent

arXiv:2511.04685v2 Announce Type: replace Abstract: In this work, we present the solution approach for the Integrated Healthcare Timetabling Competition 2024 submitted by Team Twente, which ultimately

model-releasesarxiv-cs-ai
5 May 2026
Model Releases

A Light Weight Multi-Features-View Convolution Neural Network For Plant Disease Identification

DGX agent

arXiv:2605.00903v1 Announce Type: new Abstract: Agriculture is a key sector of the economies of developing countries. It serves as a primary source of income and employment for rural populations. Howe

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

A multilingual hallucination benchmark: MultiWikiQHalluA

DGX agent

arXiv:2605.02504v1 Announce Type: new Abstract: Most hallucination evaluations focus on English, leaving it unclear whether findings transfer to lower-resource languages. We investigate faithfulness h

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

A Parameter-Free First-Order Algorithm for Non-Convex Optimization with ilde{mkern1mu O}(epsilon^{-5/3}) Global Rate

DGX agent

arXiv:2605.02127v1 Announce Type: cross Abstract: We introduce PF-AGD, the first parameter-free, deterministic, accelerated first-order method to achieve O(epsilon^{-5/3}log(1/epsilon)) oracle complex

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

A Systematic Benchmark of Machine Transliteration Models for the Tajik-Farsi Language Pair: A Comparative Study from Rule-Based to Transformer Architectures

DGX agent

arXiv:2605.02270v1 Announce Type: new Abstract: This paper presents the first comprehensive comparative analysis of modern machine learning architectures for transliteration between Tajik (Cyrillic sc

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Accelerating battery research with an AI interface between FINALES and Kadi4Mat

DGX agent

arXiv:2605.00909v1 Announce Type: cross Abstract: The time-consuming formation process critically impacts the longevity of sodium-ion coin cells and End Of Life (EOL) performance. This study aims to o

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

Accurate Legal Reasoning at Scale: Neuro-Symbolic Offloading and Structural Auditability for Robust Legal Adjudication

DGX agent

arXiv:2605.02472v1 Announce Type: new Abstract: Legal texts often contain computational legal clauses--provisions whose understanding requires complex logic. While frontier Large Reasoning Models (LRM

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Action Agent: Agentic Video Generation Meets Flow-Constrained Diffusion

DGX agent

arXiv:2605.01477v1 Announce Type: new Abstract: We present Action Agent, a two-stage framework that unifies agentic navigation video generation with flow-constrained diffusion control for multi-embodi

model-releasesarxiv-cs-ro
5 May 2026
Model Releases

Activation Compression in LLMs: Theoretical Analysis and Efficient Algorithm

DGX agent

arXiv:2605.01255v1 Announce Type: new Abstract: Training large language models (LLMs) is highly memory-intensive, as training must store not only weights and optimizer states but also intermediate act

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

AdamO: A Collapse-Suppressed Optimizer for Offline RL

DGX agent

arXiv:2605.01968v1 Announce Type: new Abstract: Offline reinforcement learning (RL) can fail spectacularly when bootstrapped temporal-difference (TD) updates amplify their own errors, driving the crit

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

Adapting Vision-Language Foundation Model for Next Generation Medical Ultrasound Image Analysis

DGX agent

arXiv:2506.08849v4 Announce Type: replace Abstract: Vision-Language Foundation Models (VLFMs) exhibit remarkable generalization, yet their direct application to medical ultrasound is severely hindered

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

Adaptive Estimation and Inference in Semi-parametric Heterogeneous Clustered Multitask Learning via Neyman Orthogonality

DGX agent

arXiv:2605.01907v1 Announce Type: cross Abstract: We study clustered multitask learning in a semiparametric setting where tasks share a latent cluster structure in their target parameters but exhibit

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

Adaptive Texture-aware Masking for Self-Supervised Learning in 3D Dental CBCT Analysis

DGX agent

arXiv:2605.01741v1 Announce Type: new Abstract: Cone Beam Computed Tomography (CBCT) is pivotal for 3D diagnostic imaging in dentistry. However, the development of robust AI models for volumetric anal

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

Addressing Data Scarcity in Bangla Fake News Detection: An LLM-Based Dataset Augmentation Approach

DGX agent

arXiv:2605.01292v1 Announce Type: new Abstract: The growing spread of misinformation in digital media highlights the need for reliable fake news detection systems, yet progress in under-resourced lang

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Adoption and Use of LLMs at an Academic Medical Center

DGX agent

arXiv:2602.00074v2 Announce Type: replace-cross Abstract: While large language models (LLMs) can support clinical documentation needs, standalone tools struggle with 'workflow friction' from manual da

model-releasesarxiv-cs-ai
5 May 2026
Model Releases

AEM: Adaptive Entropy Modulation for Multi-Turn Agentic Reinforcement Learning

DGX agent

arXiv:2605.00425v1 Announce Type: new Abstract: Reinforcement learning (RL) has significantly advanced the ability of large language model (LLM) agents to interact with environments and solve multi-tu

model-releasesarxiv-cs-ai
5 May 2026
Model Releases

Agent Factory Recap: How Gemma 4 Taught Itself Physics

DGX agent

In this episode of The Agent Factory, Vlad Kolesnikov and I sat down with Omar Sanseviero from the Developer Experience team at Google DeepMind. We explored the groundbreaking release of Gemma 4: a ne

model-releasesgoogle-cloud-ai
5 May 2026
Model Releases

Agentic AI for Trip Planning Optimization Application

DGX agent

arXiv:2605.00276v1 Announce Type: new Abstract: Trip planning for intelligent vehicles increasingly requires selecting optimal routes rather than merely producing feasible itineraries, as interacting

model-releasesarxiv-cs-ai
5 May 2026
Model Releases

Agentopic: A Generative AI Agent Workflow for Explainable Topic Modeling

DGX agent

arXiv:2605.00833v1 Announce Type: new Abstract: Agentopic is a novel agent-based workflow for explainable topic modeling that leverages the reasoning capabilities of Large Language Models (LLMs). Exis

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

AI agent startup Sierra valued at 15B in new 950M funding round

DGX agent

Eight months after closing a 350 million funding round, Sierra Technologies Inc. today announced that it has raised an additional 950 million at a 15 billion valuation. Alphabet Inc.’s GV venture capi

model-releasessiliconangle
5 May 2026
Model Releases

AI-Driven Expansion and Application of the Alexandria Database

DGX agent

arXiv:2512.09169v2 Announce Type: replace-cross Abstract: We present a novel multi-stage workflow for computational materials discovery that achieves a 99% success rate in identifying compounds within

model-releasesarxiv-cs-ai
5 May 2026
Model Releases

AI Gateway lets you route to any model. On May 13 in SF, we're hosting a builder night powered by those models. Pick one, build, demo. Audie…

DGX agent

AI Gateway lets you route to any model. On May 13 in SF, we're hosting a builder night powered by those models. Pick one, build, demo. Audience votes on best build. With @AnthropicAI, @MiniMax_AI, & @

model-releaseskimi-moonshot--x
5 May 2026
Model Releases

Ai2 releases MolmoAct 2, enhancing robot intelligence in the real world

DGX agent

Seattle-based artificial intelligence research institute Ai2, the Allen Institute for AI, today announced its next-generation open-source foundation artificial intelligence models, aimed at enabling r

model-releasessiliconangle
5 May 2026
Model Releases

Aligning LLMs with Biomedical Knowledge using Balanced Fine-Tuning

DGX agent

arXiv:2511.21075v3 Announce Type: replace Abstract: Engineering LLMs to accelerate life sciences research requires a robust alignment with biomedical knowledge. We observe that biomedical text exhibit

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

An ALE-Consistent Graph Neural Operator-Transformer Framework for Fluid-Structure Interaction

DGX agent

arXiv:2605.00937v1 Announce Type: cross Abstract: We propose an arbitrary Lagrangian-Eulerian (ALE)-consistent machine learning framework for long-term fluid-structure interaction (FSI) prediction on

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

An Efficient Metric for Data Quality Measurement in Imitation Learning

DGX agent

arXiv:2605.01544v1 Announce Type: new Abstract: Imitation learning (IL) has seen remarkable progress, yet field deployment of IL-powered robots remains hindered by the challenge of out-of-distribution

model-releasesarxiv-cs-ro
5 May 2026
Model Releases

AnchorD: Metric Grounding of Monocular Depth Using Factor Graphs

DGX agent

arXiv:2605.02667v1 Announce Type: cross Abstract: Dense and accurate depth estimation is essential for robotic manipulation, grasping, and navigation, yet currently available depth sensors are prone t

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

ARA: Agentic Reproducibility Assessment For Scalable Support Of Scientific Peer-Review

DGX agent

arXiv:2605.02651v1 Announce Type: cross Abstract: Scientific peer review increasingly struggles to assess reproducibility at the scale and complexity of modern research output. Evaluating reproducibil

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

ARIS: Agentic and Relationship Intelligence System for Social Robots

DGX agent

arXiv:2605.00943v1 Announce Type: new Abstract: Foundational models have advanced social robotics, enabling richer perception and communicative interaction with users. However, current systems still s

model-releasesarxiv-cs-ro
5 May 2026
Model Releases

Arithmetic in the Wild: Llama uses Base-10 Addition to Reason About Cyclic Concepts

DGX agent

arXiv:2605.01148v1 Announce Type: cross Abstract: Does structure in representations imply structure in computation? We study how Llama-3.1-8B reasons over cyclic concepts (e.g., 'what month is six mon

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

ARMOR 2025: A Military-Aligned Benchmark for Evaluating Large Language Model Safety Beyond Civilian Contexts

DGX agent

arXiv:2605.00245v1 Announce Type: new Abstract: Large language models (LLMs) are now being explored for defense applications that require reliable and legally compliant decision support. They also hol

model-releasesarxiv-cs-ai
5 May 2026
Model Releases

Assistance Without Interruption: A Benchmark and LLM-based Framework for Non-Intrusive Human-Robot Assistance

DGX agent

arXiv:2605.01368v1 Announce Type: new Abstract: Human-robot interaction (HRI) has long studied how agents and people coordinate to achieve shared goals. In this work, we formalize and benchmark the no

model-releasesarxiv-cs-ro
5 May 2026
Model Releases

ATR-Bench: A Federated Learning Benchmark for Adaptation, Trust, and Reasoning

DGX agent

arXiv:2505.16850v2 Announce Type: replace-cross Abstract: Federated Learning (FL) has emerged as a promising paradigm for collaborative model training while preserving data privacy across decentralize

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Attention Is Where You Attack

DGX agent

arXiv:2605.00236v1 Announce Type: cross Abstract: Safety-aligned large language models rely on RLHF and instruction tuning to refuse harmful requests, yet the internal mechanisms implementing safety b

model-releasesarxiv-cs-ai
5 May 2026
Model Releases

AttnRouter: Per-Category Attention Routing for Training-Free Image Editing on MMDiT

DGX agent

arXiv:2605.01480v1 Announce Type: new Abstract: We study training-free image editing on Qwen-Image-Edit-2511, a 60-block multi-modal diffusion transformer (MMDiT) that concatenates noise and source-im

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

Automated Interpretability and Feature Discovery in Language Models with Agents

DGX agent

arXiv:2605.01555v1 Announce Type: new Abstract: We introduce an autonomous multiagent framework for mechanistic interpretability that automates both explaining and finding internal features in large l

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

AutoSpatial: Visual-Language Reasoning for Social Robot Navigation through Efficient Spatial Reasoning Learning

DGX agent

arXiv:2503.07557v2 Announce Type: replace Abstract: We present a novel method, AutoSpatial, an efficient approach with structured spatial grounding to enhance VLMs' spatial reasoning. By combining min

model-releasesarxiv-cs-ro
5 May 2026
← Previous
1…356357358359360…471
Next →