AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,520 results
30 Apr 2026

The Hermes Agent Creative Hackathon ends on Sunday, just enough time for a last minute weekend project! Here's a thread of some of the power…

Model ReleasesDGX agent

The Hermes Agent Creative Hackathon ends on Sunday, just enough time for a last minute weekend project! Here's a thread of some of the powerful creative skills and tools we've released that might help

The model wars are over. Now, Google is fighting for something bigger

Model ReleasesDGX agent

Whoever controls the agentic control plane controls enterprise AI. Google LLC just showed up to that fight with everything it has. The company claiming it can own the full stack arrived at Google Clou

The Prompt Engineering Report Distilled: Quick Start Guide for Life Sciences

Model Releases

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2509.11295v2 Announce Type: replace Abstract: Developing effective prompts demands significant cognitive investment to generate reliable, high-quality responses from Large Language Models (LLMs)

The Unseen Adversaries: Robust and Generalized Defense Against Adversarial Patches

Model ReleasesDGX agent

arXiv:2604.26317v1 Announce Type: new Abstract: The vulnerabilities of deep neural networks against singularities have raised serious concerns regarding their deployment in the physical world. One of

Theory-Grounded Evaluation Exposes the Authorship Gap in LLM Personalization

Model ReleasesDGX agent

arXiv:2604.26460v1 Announce Type: new Abstract: Stylistic personalization - making LLMs write in a specific individual's style, rather than merely adapting to task preferences - lacks evaluation groun

Thinking with Drafting: Optical Decompression via Logical Reconstruction

Model ReleasesDGX agent

arXiv:2602.11731v2 Announce Type: replace Abstract: Existing multimodal large language models have achieved high-fidelity visual perception and exploratory visual generation. However, a precision para

This is really well thought out. Filesystems are the new default abstraction for agents to interact with documents (the new RAG stack in 202…

Model ReleasesDGX agent

This is really well thought out. Filesystems are the new default abstraction for agents to interact with documents (the new RAG stack in 2026). The issue is actually figuring out how to productize thi

This startup’s new mechanistic interpretability tool lets you debug LLMs

Model ReleasesDGX agent

The San Francisco–based startup Goodfire just released a new tool, called Silico, that lets researchers and engineers peer inside an AI model and adjust its parameters—the settings that determine a mo

TildeOpen LLM: Leveraging Curriculum Learning to Achieve Equitable Language Representation

Model ReleasesDGX agent

arXiv:2603.08182v2 Announce Type: replace-cross Abstract: Large language models often underperform in many European languages due to the dominance of English and a few high-resource languages in train

Time Blindness: Why Video-Language Models Can't See What Humans Can?

Model ReleasesDGX agent

arXiv:2505.24867v2 Announce Type: replace-cross Abstract: Recent advances in vision-language models (VLMs) have made impressive strides in understanding spatio-temporal relationships in videos. Howeve

Time series classification with random convolution kernels: pooling operators and input representations matter

Model ReleasesDGX agent

arXiv:2409.01115v5 Announce Type: replace Abstract: This article presents a new approach based on MiniRocket, called SelF-Rocket, for fast time series classification (TSC). Unlike existing approaches

TinyR1-32B-Preview: Boosting Accuracy with Branch-Merge Distillation

Model ReleasesDGX agent

arXiv:2503.04872v3 Announce Type: replace-cross Abstract: The challenge of reducing the size of Large Language Models (LLMs) while maintaining their performance has gained significant attention. Howev

Today we’re releasing Qwen-Scope 🔭, an open suite of sparse autoencoders for the Qwen model family. It turns SAE features into practical to…

Model ReleasesDGX agent

Today we’re releasing Qwen-Scope 🔭, an open suite of sparse autoencoders for the Qwen model family. It turns SAE features into practical tools: 🎯 Inference — Steer model outputs by directly manipulati

Train a TensorFlow object detection model – then deploy it on a robot 🤖 Iulia Feroli (@iuliaferoli) shows how to turn a notebook into a rea…

Model ReleasesDGX agent

Train a TensorFlow object detection model – then deploy it on a robot 🤖 Iulia Feroli (@iuliaferoli) shows how to turn a notebook into a real-time object detection app. This tutorial works for any proj

Training-Free Adaptation of New-Generation LLMs using Legacy Clinical Models

Model ReleasesDGX agent

arXiv:2601.03423v3 Announce Type: replace-cross Abstract: Adapting language models to the clinical domain through continued pretraining and instruction tuning requires costly retraining for each new m

Training-Free Loosely Speculative Decoding: Accepting Semantically Correct Drafts Beyond Exact Match

Model ReleasesDGX agent

arXiv:2511.22972v3 Announce Type: replace Abstract: Large language models (LLMs) achieve strong performance across diverse tasks but suffer from high inference latency due to their autoregressive gene

tweeted about this yesterday and Cursor already dropped the alpha today! 🚀very cool to see how us, them, and others have converged on good …

Model ReleasesDGX agent

tweeted about this yesterday and Cursor already dropped the alpha today! 🚀very cool to see how us, them, and others have converged on good design patterns in Agent + Harness Engineering: 1. Tuning dif

Value-Guided Iterative Refinement and the DIQ-H Benchmark for Evaluating VLM Robustness

Model ReleasesDGX agent

arXiv:2512.03992v2 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) are essential for embodied AI and safety-critical applications, such as robotics and autonomous systems. However

VIGNETTE: Socially Grounded Bias Evaluation for Vision-Language Models

Model ReleasesDGX agent

arXiv:2505.22897v2 Announce Type: replace Abstract: While bias in large language models (LLMs) is well-studied, similar concerns in vision-language models (VLMs) have received comparatively less atten

VLN-Cache: Enabling Token Caching for VLN Models with Visual/Semantic Dynamics Awareness

Model ReleasesDGX agent

arXiv:2603.07080v3 Announce Type: replace-cross Abstract: Vision-and-Language Navigation (VLN) increasingly relies on large vision-language models, but their inference cost conflicts with real-time de

VulStyle: A Multi-Modal Pre-Training for Code Stylometry-Augmented Vulnerability Detection

Model ReleasesDGX agent

arXiv:2604.26313v1 Announce Type: cross Abstract: We present VulStyle, a multi-modal software vulnerability detection model that jointly encodes function-level source code, non-terminal Abstract Synta

We need RSS for sharing abundant vibe-coded apps

Model ReleasesDGX agent

We need RSS for sharing abundant vibe-coded apps Matt Webb: I would love an RSS web feed for all those various tools and apps pages, each item with an “Install” button. (But install to where?) The les

WebAggregator: Enhancing Compositional Reasoning Capabilities of Deep Research Agent Foundation Models

Model ReleasesDGX agent

arXiv:2510.14438v2 Announce Type: replace Abstract: The hallmark of Deep Research agents lies in compositional reasoning, the capacity to aggregate distributed, heterogeneous information into coherent

we're starting rollout of GPT-5.5-Cyber, a frontier cybersecurity model, to critical cyber defenders in the next few days. we will work with…

Model ReleasesDGX agent

we're starting rollout of GPT-5.5-Cyber, a frontier cybersecurity model, to critical cyber defenders in the next few days. we will work with the entire ecosystem and the government to figure out trust

We've partnered with @OpenAI to offer GPT-5.5 in Devin at 50% off through May 14 starting today.

Model ReleasesDGX agent

We've partnered with @OpenAI to offer GPT-5.5 in Devin at 50% off through May 14 starting today. GPT-5.5 is now available in Devin as an Agent Preview! GPT-5.5 has set a new bar for what's possible wi

We've partnered with @OpenAI to offer GPT-5.5 in Windsurf at 50% off through May 14 starting today.

Model ReleasesDGX agent

Windsurf has announced a partnership with OpenAI to provide GPT-5.5 access within their platform at a 50% discount through May 14. This promotional offer began on the date of the announcement and appe

What Google Cloud announced in AI this month

Model ReleasesDGX agent

Editor’s note: Want to keep up with the latest from Google Cloud? Check back here for a monthly recap of our latest updates, announcements, resources, events, learning opportunities, and more. We host

What people call 'distillation' is a super common practice (you use other models to benchmark your model, to evaluate your inputs or to add …

Model ReleasesDGX agent

What people call 'distillation' is a super common practice (you use other models to benchmark your model, to evaluate your inputs or to add a little bit to your datasets) that in my opinion should be

When to Retrieve During Reasoning: Adaptive Retrieval for Large Reasoning Models

Model ReleasesDGX agent

arXiv:2604.26649v1 Announce Type: cross Abstract: Large reasoning models such as DeepSeek-R1 and OpenAI o1 generate extended chains of thought spanning thousands of tokens, yet their integration with

Wordle 1,775 4/6 ⬛⬛🟨🟨⬛ 🟨⬛⬛⬛⬛ ⬛🟨🟨⬛🟩 🟩🟩🟩🟩🟩

Model ReleasesDGX agent

This post documents a Wordle game result where the player solved puzzle #1,775 in 4 attempts, using the color-coded feedback system (gray for incorrect letters, yellow for correct letters in wrong pos

Work faster with Codex. https://chatgpt.com/codex/for-work/

Model ReleasesDGX agent

Codex is OpenAI's AI tool designed to accelerate work productivity by generating and understanding code, enabling developers to write, debug, and complete programming tasks more efficiently. The resou

29 Apr 2026

$3/million output tokens. Qwen 3.5 Plus is basically a frontier model. Let that sink in.

Model ReleasesDGX agent

$3/million output tokens. Qwen 3.5 Plus is basically a frontier model. Let that sink in. Introducing Qwen3.6-Plus from @Alibaba_Qwen, a 1M-context model built for real-world agents, agentic coding, an

A Comparative Analysis on the Performance of Upper Confidence Bound Algorithms in Adaptive Deep Neural Networks

Model ReleasesDGX agent

arXiv:2604.24810v1 Announce Type: new Abstract: Edge computing environments impose strict constraints on energy consumption and latency, making the deployment of deep neural networks a significant cha

A Comparative Study in Surgical AI: Datasets, Foundation Models, and Barriers to Med-AGI

Model ReleasesDGX agent

arXiv:2603.27341v2 Announce Type: replace-cross Abstract: Recent Artificial Intelligence (AI) models have matched or exceeded human experts in several benchmarks of biomedical task performance, but su

A million-token context window is not a strategy. 🛑 Our Head of DevRel, @RoieSchwabco , explains why dumping data is killing your RAG perfo…

Model ReleasesDGX agent

A million-token context window is not a strategy. 🛑 Our Head of DevRel, @RoieSchwabco , explains why dumping data is killing your RAG performance: 📍 One needle in a haystack? Easy. 📍 Multiple needles?

Adaptable phase retrieval for coherent transition radiation spectroscopy based on differentiable physics information

Model ReleasesDGX agent

arXiv:2604.25489v1 Announce Type: cross Abstract: Coherent transition radiation (CTR) spectroscopy is a critical diagnostic for characterizing the longitudinal structure of relativistic electron bunch

AdaTooler-V: Adaptive Tool-Use for Images and Videos

Model ReleasesDGX agent

arXiv:2512.16918v3 Announce Type: replace Abstract: Recent advances have shown that multimodal large language models (MLLMs) benefit from multimodal interleaved chain-of-thought (CoT) with vision tool

ADE: Adaptive Dictionary Embeddings -- Scaling Multi-Anchor Representations to Large Language Models

Model ReleasesDGX agent

arXiv:2604.24940v1 Announce Type: new Abstract: Word embeddings are fundamental to natural language processing, yet traditional approaches represent each word with a single vector, creating representa

Agent-Diff: Benchmarking LLM Agents on Enterprise API Tasks via Code Execution with State-Diff-Based Evaluation

Model ReleasesDGX agent

arXiv:2602.11224v3 Announce Type: replace-cross Abstract: We present Agent-Diff, a novel benchmarking framework for evaluating agentic Large Language Models (LLMs) on real-world productivity software

Agentic Harness Engineering: Observability-Driven Automatic Evolution of Coding-Agent Harnesses

Model ReleasesDGX agent

arXiv:2604.25850v1 Announce Type: new Abstract: Harnesses have become a central determinant of coding-agent performance, shaping how models interact with repositories, tools, and execution environment

Align then Adapt: Rethinking Parameter-Efficient Transfer Learning in 4D Perception

Model ReleasesDGX agent

arXiv:2602.23069v2 Announce Type: replace Abstract: Point cloud video understanding is critical for robotics as it accurately encodes motion and scene interaction. We recognize that 4D datasets are fa

Am I the only one that still likes ChatGPT? And I use Claude also

Model ReleasesDGX agent

A Reddit discussion from r/ChatGPT in which a user expresses their continued preference for ChatGPT while also using Claude, likely exploring whether other users share similar sentiments about ChatGPT

An Investigation of Linguistic Biases in LLM-Based Recommendations

Model ReleasesDGX agent

arXiv:2604.25456v1 Announce Type: new Abstract: We investigate linguistic biases in LLM-based restaurant and product recommendations given prompts varying across Southern American English (AE), Indian

Analyzing LLM Reasoning to Uncover Mental Health Stigma

Model ReleasesDGX agent

arXiv:2604.25053v1 Announce Type: new Abstract: While large language models (LLMs) are increasingly being explored for mental health applications, recent studies reveal that they can exhibit stigma to

Application of a Mixture of Experts-based Foundation Model to the GlueX DIRC Detector

Model ReleasesDGX agent

arXiv:2604.24775v1 Announce Type: cross Abstract: We present a Mixture-of-Experts-based foundation model applied to the GlueX DIRC detector at Jefferson Lab, demonstrating its utility as a unified fra

AQUA-Bench: Beyond Finding Answers to Knowing When There Are None in Audio Question Answering

Model ReleasesDGX agent

arXiv:2601.12248v2 Announce Type: replace-cross Abstract: Recent advances in audio-aware large language models have shown strong performance on audio question answering. However, existing benchmarks m

Architecture Determines Observability in Transformers

Model ReleasesDGX agent

arXiv:2604.24801v1 Announce Type: new Abstract: Autoregressive transformers make confident errors, but activation monitoring can catch them only if the model preserves an internal signal that output c

As models, contexts, and workloads grow, hidden assumptions in inference infrastructure can surface as output anomalies. Reliability require…

Model ReleasesDGX agent

As models, contexts, and workloads grow, hidden assumptions in inference infrastructure can surface as output anomalies. Reliability requires more than throughput, latency, and availability. It also r

asRoBallet: Closing the Sim2Real Gap via Friction-Aware Reinforcement Learning for Underactuated Spherical Dynamics

Model ReleasesDGX agent

arXiv:2604.24916v1 Announce Type: new Abstract: We introduce asRoBallet, to the best of our knowledge, the first successful deployment of reinforcement learning (RL) on a humanoid ballbot hardware. Hi

Auvik launches Aurora AI agents to speed ticket resolution and prevent outages

Model ReleasesDGX agent

Information technology management software provider Auvik Networks Inc. today announced the launch of Auvik Aurora: artificial intelligence-powered IT agents that are designed to help IT professionals

Aviatrix launches AI agent containment platform for cloud workloads

Model ReleasesDGX agent

Aviatrix Inc. today announced the launch of a new platform designed to contain artificial intelligence agents and enforce security controls and communications across AI workloads without changing AI a

Below-Chance Blindness: Prompted Underperformance in Small LLMs Produces Positional Bias Rather than Answer Avoidance

Model ReleasesDGX agent

arXiv:2604.25249v1 Announce Type: new Abstract: Detecting sandbagging--the deliberate underperformance on capability evaluations--is an open problem in AI safety. We tested whether symptom validity te

BenchGuard: Who Guards the Benchmarks? Automated Auditing of LLM Agent Benchmarks

Model ReleasesDGX agent

arXiv:2604.24955v1 Announce Type: new Abstract: As benchmarks grow in complexity, many apparent agent failures are not failures of the agent at all - they are failures of the benchmark itself: broken

Benchmarking and Adapting On-Device LLMs for Clinical Decision Support

Model ReleasesDGX agent

arXiv:2601.03266v2 Announce Type: replace Abstract: Large language models (LLMs) have rapidly advanced in clinical decision-making, yet the deployment of proprietary systems is hindered by privacy con

Benchmarking and Improving GUI Agents in High-Dynamic Environments

Model ReleasesDGX agent

arXiv:2604.25380v1 Announce Type: new Abstract: Recent advancements in Graphical User Interface (GUI) agents have predominantly focused on training paradigms like supervised fine-tuning (SFT) and rein

Benchmarking Layout-Guided Diffusion Models through Unified Semantic-Spatial Evaluation in Closed and Open Settings

Model ReleasesDGX agent

arXiv:2604.25358v1 Announce Type: new Abstract: Evaluating layout-guided text-to-image generative models requires assessing both semantic alignment with textual prompts and spatial fidelity to prescri

Benchmarking OCR Pipelines with Adaptive Enhancement for Multi-Domain Retail Bill Digitization

Model ReleasesDGX agent

arXiv:2604.25176v1 Announce Type: new Abstract: The digitization of multi-domain retail billing documents remains a challenging task due to variability in scan quality, layout heterogeneity, and domai

Beyond I'm Sorry, I Can't: Dissecting Large Language Model Refusal

Model ReleasesDGX agent

arXiv:2509.09708v3 Announce Type: replace Abstract: Refusal on harmful prompts is a key safety behaviour in instruction-tuned large language models (LLMs), yet the internal causes of this behaviour re

BifDet: A 3D Bifurcation Detection Dataset for Airway-Tree Modeling

Model ReleasesDGX agent

arXiv:2604.24999v1 Announce Type: new Abstract: Thoracic Computed Tomography (CT) scans offer detailed insights into the intricate branching network of the airway tree, which is essential for understa

BLASST: Dynamic BLocked Attention Sparsity via Softmax Thresholding

Model ReleasesDGX agent

arXiv:2512.12087v3 Announce Type: replace Abstract: The growing demand for long-context inference capabilities in Large Language Models (LLMs) has intensified the computational and memory bottlenecks

← Previous
1…299300301302303…376
Next →