AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent
84,630Total entries
1Added by human
84,629Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,611 results
28 May 2026

VULPO: Context-Aware Vulnerability Detection via On-Policy LLM Optimization

Model ReleasesDGX agent

arXiv:2511.11896v3 Announce Type: replace-cross Abstract: Large language models (LLMs) have recently shown strong potential in vulnerability detection (VD). However, accurately detecting vulnerabiliti

wait… if most people think 5.5 is better than 4.7, i assume that’s due to terminal coding benchmark… 4.8 is still outperformed by 5.5

Model ReleasesDGX agent

wait… if most people think 5.5 is better than 4.7, i assume that’s due to terminal coding benchmark… 4.8 is still outperformed by 5.5 Introducing Claude Opus 4.8: it builds on Opus 4.7 with sharper ju

We also shipped dynamic workflows in Claude Code (research preview), for tasks too big for one pass. Make sure to default to auto mode so Cl…


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model ReleasesDGX agent

We also shipped dynamic workflows in Claude Code (research preview), for tasks too big for one pass. Make sure to default to auto mode so Claude isn't stopping for permissions. It's token-intensive, s

Weak Convergence Analysis of Online Neural Actor-Critic Algorithms

Model ReleasesDGX agent

arXiv:2403.16825v2 Announce Type: replace Abstract: We prove that a single-layer neural network trained with the online actor critic algorithm converges in distribution to a random ordinary differenti

WeatherCity: Urban Scene Reconstruction with Controllable Multi-Weather Transformation

Model ReleasesDGX agent

arXiv:2602.22096v2 Announce Type: replace Abstract: Editable high-fidelity 4D scenes are crucial for autonomous driving, as they can be applied to end-to-end training and closed-loop simulation. Howev

We're releasing Paris 2.0, which, to our knowledge, is the world's first decentralized trained video generation model. We benchmarked it aga…

Model ReleasesDGX agent

We're releasing Paris 2.0, which, to our knowledge, is the world's first decentralized trained video generation model. We benchmarked it against a monolithic model trained on the same data and compute

We've raised 65 billion in Series H funding at a 965 billion post-money valuation, led by @AltimeterCap, Dragoneer, @Greenoaks, and @sequo…

Model ReleasesDGX agent

We've raised 65 billion in Series H funding at a 965 billion post-money valuation, led by @AltimeterCap, Dragoneer, @Greenoaks, and @sequoia. This investment will help us advance our research and expa

What-If World: A Causal Benchmark for General World Models in Embodied Scenarios

Model ReleasesDGX agent

arXiv:2605.27589v1 Announce Type: new Abstract: Video generation models are increasingly used as world simulators for tasks like driving and robotic manipulation. What matters in these settings is not

When Context Flips, Safety Breaks: Diagnosing Brittle Safety in Aligned Language Models

Model ReleasesDGX agent

arXiv:2605.27851v1 Announce Type: new Abstract: Safety benchmark scores provide incomplete evidence of deployment readiness: aligned language models often adhere to rigid rules even when a situational

When do complex-valued neural networks help? A study of representation, geometry, and optimization

Model ReleasesDGX agent

arXiv:2605.27673v1 Announce Type: new Abstract: Complex-valued Neural Networks (CVNNs) are often motivated by domains where information is naturally encoded in magnitude and phase. Yet complex-valued

When Interpretability Is Unequally Distributed: Fairness in Hybrid Interpretable Models

Model ReleasesDGX agent

arXiv:2605.28626v1 Announce Type: new Abstract: Hybrid interpretable models combine a transparent component with a black-box model by assigning some examples to the former and deferring the rest to th

Who's going first

Model ReleasesDGX agent

Who's going first Introducing Claude Opus 4.8: it builds on Opus 4.7 with sharper judgment, more honesty about its own progress, and the ability to work independently for longer than its predecessors.

Whose Name Comes Up? III: Persona Prompting Effects in LLM-Based Scholar Recommendation

Model ReleasesDGX agent

arXiv:2605.28187v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used as scholar recommenders, shaping who is seen as an expert in academia. Existing audits remain Engli

Why LLMs Fail at Causal Discovery and How Interventional Agents Escape

Model ReleasesDGX agent

arXiv:2605.27567v1 Announce Type: new Abstract: Causal discovery is a cornerstone of scientific reasoning, yet whether large language models can perform it reliably remains an open question. Recent be

Wordle 1,803 5/6 🟩🟩⬛⬛⬛ ⬛⬛⬛⬛⬛ 🟩🟩🟩⬛⬛ 🟩🟩🟩⬛⬛ 🟩🟩🟩🟩🟩

Model ReleasesDGX agent

This entry documents a Wordle game result (puzzle #1,803) where the player achieved a solution in 5 out of 6 allowed guesses. The color-coded emoji sequence shows the player's guess progression, with

XTransfer: Modality-Agnostic Few-Shot Model Transfer for Human Sensing at the Edge

Model ReleasesDGX agent

arXiv:2506.22726v4 Announce Type: replace Abstract: Deep learning for human sensing on edge systems presents significant potential for smart applications. However, its training and development are hin

You Live More Than Once: Towards Hierarchical Skill Meta-Evolving

Model ReleasesDGX agent

arXiv:2605.28390v1 Announce Type: new Abstract: Test-time skill evolving is regarded as a new paradigm for enhancing deployed agentic systems. Existing works mainly focus on hard-coded skill evolving

You Only Align Once: Propagating Cooperative Behaviors in Multi-Agent Systems through Seed Agents

Model ReleasesDGX agent

arXiv:2605.27586v1 Announce Type: cross Abstract: Ensuring agent behaviors in distributed open multi-agent systems remains challenging, especially as populations grow and unaligned agents may exist. W

ZipRL: Adaptive Multi-Turn Context Compression with Hindsight Response Replay

Model ReleasesDGX agent

arXiv:2605.28069v1 Announce Type: new Abstract: Adaptive context compression is vital for scaling Large Language Models (LLMs) to complex, multi-turn agent tasks. However, rule-based compression metho

27 May 2026

A closer look at what we released today 🧵 ESMC is a language model trained on billions of protein sequences spanning the full diversity of …

Model ReleasesDGX agent

A closer look at what we released today 🧵 ESMC is a language model trained on billions of protein sequences spanning the full diversity of life. Trained across 2.8 billion sequences, the model is expo

A Dataset of Robot-Patient and Doctor-Patient Medical Dialogues for Spoken Language Processing Tasks

Model ReleasesDGX agent

arXiv:2605.26747v1 Announce Type: new Abstract: Large Language Models (LLMs) have brought huge improvements to Artificial Intelligence (AI), which can be applied to general-purpose tasks. However, the

A deep dive into how Anthropic's Claude Code and Peter Steinberger's OpenClaw unleashed the AI agent revolution that is rapidly transforming modern computing (Steven Levy/Wired)

Model ReleasesDGX agent

Steven Levy / Wired: A deep dive into how Anthropic's Claude Code and Peter Steinberger's OpenClaw unleashed the AI agent revolution that is rapidly transforming modern computing — The definitive stor

A Deep State-Space Model Compression Method using Upper Bound on Output Error

Model ReleasesDGX agent

arXiv:2510.14542v2 Announce Type: replace-cross Abstract: We study deep state-space models (Deep SSMs) that contain linear quadratic-output (LQO) systems as internal blocks and present a compression m

A Guide to AI Cold Starts on Cloud Run

Model ReleasesDGX agent

I saw a developer asking on Reddit if there was any “sane way” to manage Cloud Run cold starts for AI across multiple regions. They were experiencing startup latencies of up to 20 seconds, a frustrati

A Hybrid Vision-Language Architecture for Automated Defect Reasoning and Report Generation in Industrial Inspection

Model ReleasesDGX agent

arXiv:2605.26533v1 Announce Type: cross Abstract: Automated industrial inspection requires both precise defect localization and structured maintenance report generation; in current practice these task

A newly released AI tool has generated an atlas of more than one billion predicted protein structures and billions more protein sequences. h…

Model ReleasesDGX agent

DeepMind's AlphaFold3 and related tools have generated a comprehensive atlas containing over one billion predicted protein structures and additional billions of protein sequences, representing a major

AdaSD: Adaptive Speculative Decoding for Efficient Language Model Inference

Model ReleasesDGX agent

arXiv:2512.11280v2 Announce Type: replace Abstract: Large language models (LLMs) have achieved remarkable performance across a wide range of tasks, but their increasing parameter sizes significantly s

ADRD-Bench: A Preliminary LLM Benchmark for Alzheimer's Disease and Related Dementias

Model ReleasesDGX agent

arXiv:2602.11460v2 Announce Type: replace Abstract: Large language models (LLMs) have shown great potential for healthcare applications. However, existing evaluation benchmarks provide minimal coverag

Advancing Creative Physical Intelligence in Large Multimodal Models

Model ReleasesDGX agent

arXiv:2605.26396v1 Announce Type: new Abstract: Large multimodal models (LMMs) have rapidly advanced in perception and reasoning; however, it remains unclear whether these capabilities generalize to d

AgentSociety: Incentivizing Agentic Social Intelligence

Model ReleasesDGX agent

arXiv:2605.26203v1 Announce Type: cross Abstract: The success of deployed agents relies on their ability to handle open-ended user requests using their inherent capabilities, not only in solving reque

AGORA: Adapter-Grounded Observation-Action Retention for Inference-Free Prompt Compression in LLM Agents

Model ReleasesDGX agent

arXiv:2605.26596v1 Announce Type: new Abstract: The token-level extractive compressors widely used for general LM context are structurally inappropriate for LLM agents: across 17 (env, backbone, metho

AI evaluation may bias perceptions: The importance of context in interpreting academic writing

Model ReleasesDGX agent

arXiv:2605.26662v1 Announce Type: cross Abstract: This paper examines how estimates of AI use in scientific writing can be biased when evaluation methods ignore contextual differences across countries

AlbanianLLMSafety: A Safety Evaluation Dataset for Large Language Models in Albanian

Model ReleasesDGX agent

arXiv:2605.26954v1 Announce Type: new Abstract: Safety evaluation of Large Language Models (LLMs) has largely focused on high-resource languages, leaving low-resource languages critically underserved.

Amazon MGM Studios announces the GenAI Creators' Fund, greenlights three AI animated series for Prime Video, and launches an AI production platform with AWS (Todd Spangler/Variety)

Model ReleasesDGX agent

Todd Spangler / Variety: Amazon MGM Studios announces the GenAI Creators' Fund, greenlights three AI animated series for Prime Video, and launches an AI production platform with AWS — Amazon MGM Studi

An End-to-End Learning Approach for Solving Capacitated Location-Routing Problems

Model ReleasesDGX agent

arXiv:2511.02525v2 Announce Type: replace-cross Abstract: The capacitated location-routing problems (CLRPs) are classical problems in combinatorial optimization, which require simultaneously making lo

An uncertainty-aware Bayesian framework for machine learning classification models: A case study in land cover classification

Model ReleasesDGX agent

arXiv:2503.21510v3 Announce Type: replace-cross Abstract: Ensuring that predictions of machine learning (ML) classification models are accompanied by uncertainty estimates is one of the main pillars o

Anchor: Mitigating Artifact Drift in Agent Benchmark Generation

Model ReleasesDGX agent

arXiv:2605.26321v1 Announce Type: new Abstract: AI agents are beginning to complete valuable, long-horizon business operations tasks, but training and evaluation environments for enterprise work still

Announcing ESMFold2, our new state-of-the-art structure prediction model capable of predicting structure from single sequences or MSAs. ESMF…

Model ReleasesDGX agent

Announcing ESMFold2, our new state-of-the-art structure prediction model capable of predicting structure from single sequences or MSAs. ESMFold2 improves on benchmarks of protein-protein interaction a

ARBITER: Reasoning Trajectory Basins and Majority Vote Failures in Test-Time Sampling

Model ReleasesDGX agent

arXiv:2605.26172v1 Announce Type: new Abstract: When language models use test-time sampling, they generate multiple reasoning trajectories and select an answer by majority vote. We show that these tra

Are Video Models Zero-Shot Learners and Reasoners in Education? EduVideoBench, A Knowledge-Skills-Attitude Benchmark for Educational Video Generation

Model ReleasesDGX agent

arXiv:2605.26918v1 Announce Type: new Abstract: Video generation models (VGMs) are rapidly entering classrooms, yet existing benchmarks evaluate only perceptual quality, intrinsic faithfulness, generi

Attribute-Based Diagnosis of LLM Alignment with Hate Speech Annotations

Model ReleasesDGX agent

arXiv:2605.27025v1 Announce Type: new Abstract: Hate speech annotation is costly, subjective, and prone to annotator disagreement, making large-scale dataset construction challenging. We systematicall

AutoDFT: A Closed-Loop Multi-Agent Framework for Autonomous DFT Calculations

Model ReleasesDGX agent

arXiv:2605.26179v1 Announce Type: cross Abstract: Density functional theory (DFT) serves as the basis for computational discovery in materials science and chemistry, yet each calculation demands exten

AWS launches Agentic Shopping Assistant to help retailers build AI tools

Model ReleasesDGX agent

Amazon Web Services Inc. today introduced a new offering designed to help retailers integrate artificial intelligence features into their online stores. AWS Agentic Shopping Assistant, or ASA, combine

Axial-Centric Cross-Plane Attention for 3D Medical Image Classification

Model ReleasesDGX agent

arXiv:2602.21636v2 Announce Type: replace Abstract: Abridged: Clinicians commonly interpret 3D medical images by examining multiple anatomical planes rather than relying on volumetric views. In clinic

BEAT: Rhythm-Elastic Alignment for Agentic Music-guided Movie Trailer Generation

Model ReleasesDGX agent

arXiv:2605.27067v1 Announce Type: new Abstract: Automatic movie trailer generation must select shots from a full-length film and synchronize them with background music. Existing methods either relegat

Benchmark Leakage Trap: Can We Trust LLM-based Recommendation?

Model ReleasesDGX agent

arXiv:2602.13626v3 Announce Type: replace Abstract: The expanding integration of Large Language Models (LLMs) into recommender systems poses critical challenges to evaluation reliability. This paper i

Benchmarking Convolutional, Transformer, Hybrid, and Vision Language Models for Multi Disease Retinal Screening

Model ReleasesDGX agent

arXiv:2605.26283v1 Announce Type: new Abstract: Modern deep learning offers powerful tools for automated retinal screening, but it remains unclear how different visual model families compare in realis

BESPOKE: Benchmark for Search-Augmented Large Language Model Personalization via Diagnostic Feedback

Model ReleasesDGX agent

arXiv:2509.21106v2 Announce Type: replace Abstract: Search-augmented large language models (LLMs) have advanced information-seeking tasks by integrating retrieval into generation, reducing users' cogn

Beyond a Single Direction: Chain-of-Thought Disrupts Simple Steering of Refusal

Model ReleasesDGX agent

arXiv:2605.26772v1 Announce Type: new Abstract: Large reasoning models (LRMs) generate chain-of-thought (CoT) traces before producing final outputs, introducing a dynamic internal state that may compl

Beyond Holistic Models: Systematic Component-level Benchmarking of Deep Multivariate Time-Series Forecasting

Model ReleasesDGX agent

arXiv:2605.26562v1 Announce Type: new Abstract: While previous research in multivariate time series forecasting has focused on developing complex holistic models, this work advocates for a shift towar

Beyond Questions: Evaluating What Large Language Models (Actually) Know

Model ReleasesDGX agent

arXiv:2605.26937v1 Announce Type: cross Abstract: Parametric knowledge in large language models (LLMs) is a cornerstone of their success, yet remains poorly understood. Existing knowledge benchmarks t

Beyond Transfer Accuracy: Faithful Circuits for Controlled Low-Resource Adaptation

Model ReleasesDGX agent

arXiv:2601.08146v3 Announce Type: replace-cross Abstract: Existing circuit discovery methods rely on templated tasks with clean counterfactuals, limiting their use on diverse natural text. We adapt Co

BeyondSWE: Can Current Code Agent Survive Beyond Single-Repo Bug Fixing?

Model ReleasesDGX agent

arXiv:2603.03194v2 Announce Type: replace Abstract: Current code-agent benchmarks primarily evaluate localized issue resolution within a single target repository, leaving under-tested many software en

BhashaSetu: A Data-Centric Approach to Low-Resource Machine Translation

Model ReleasesDGX agent

arXiv:2605.27050v1 Announce Type: new Abstract: We present BhashaSetu, a linguistically enriched English--Marathi parallel dataset addressing persistent data limitations in low-resource neural machine

Black-box Membership Inference Attacks on the Pre-training Data of Image-generation Models

Model ReleasesDGX agent

arXiv:2605.27020v1 Announce Type: cross Abstract: The rapid advancement of diffusion-based image generation models has raised serious concerns regarding potential copyright and privacy infringements i

Bridging Classification and Reconstruction: Cooperative Time Series Anomaly Detection

Model ReleasesDGX agent

arXiv:2605.26193v1 Announce Type: cross Abstract: Time series anomaly detection (TSAD) has long been a hot research topic in data mining due to its various applications. Recent studies challenge the e

Cast a Wider Net: Coordinated Pass@K Policy Optimization for Code Reasoning

Model ReleasesDGX agent

arXiv:2605.27000v1 Announce Type: cross Abstract: Repeated sampling with a verifier is the standard way to allocate test-time compute for code generation, with pass@K as the canonical metric. Yet the

Causal Representation Learning for Generalisable Recommendation

Model ReleasesDGX agent

arXiv:2605.27043v1 Announce Type: cross Abstract: Predictive models trained on observational data often fail to generalise to the distributions they encounter when deployed, especially when the traini

Cesarean Scar Defect Segmentation in Transvaginal Ultrasound Images: a Dataset and Benchmark

Model ReleasesDGX agent

arXiv:2605.26774v1 Announce Type: new Abstract: Cesarean Scar Defect (CSD) is one of the most prevalent complications following cesarean delivery. Transvaginal ultrasonography is widely used for prima

Chain Of Thought Compression: A Theoretical Analysis

Model ReleasesDGX agent

arXiv:2601.21576v2 Announce Type: replace Abstract: Chain-of-Thought (CoT) has unlocked advanced reasoning abilities of Large Language Models (LLMs) with intermediate steps, yet incurs prohibitive com

← Previous
1…204205206207208…377
Next →