AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,284 results
16 Apr 2026

ROBOGATE: Adaptive Failure Discovery for Safe Robot Policy Deployment via Two-Stage Boundary-Focused Sampling

Model ReleasesDGX agent

arXiv:2603.22126v3 Announce Type: replace Abstract: Deploying learned robot manipulation policies in industrial settings requires rigorous pre-deployment validation, yet exhaustive testing across high

Robust Low-Rank Tensor Completion based on M-product with Weighted Correlated Total Variation and Sparse Regularization

Model ReleasesDGX agent

arXiv:2604.13525v1 Announce Type: cross Abstract: The robust low-rank tensor completion problem addresses the challenge of recovering corrupted high-dimensional tensor data with missing entries, outli

Robust Reward Modeling for Large Language Models via Causal Decomposition


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases
DGX agent

arXiv:2604.13833v1 Announce Type: new Abstract: Reward models are central to aligning large language models, yet they often overfit to spurious cues such as response length and overly agreeable tone.

ROSE: Retrieval-Oriented Segmentation Enhancement

Model ReleasesDGX agent

arXiv:2604.14147v1 Announce Type: new Abstract: Existing segmentation models based on multimodal large language models (MLLMs), such as LISA, often struggle with novel or emerging entities due to thei

RPS: Information Elicitation with Reinforcement Prompt Selection

Model ReleasesDGX agent

arXiv:2604.13817v1 Announce Type: new Abstract: Large language models (LLMs) have shown remarkable capabilities in dialogue generation and reasoning, yet their effectiveness in eliciting user-known bu

Scaling Test-Time Compute to Achieve IOI Gold Medal with Open-Weight Models

Model ReleasesDGX agent

arXiv:2510.14232v2 Announce Type: replace-cross Abstract: Competitive programming has become a rigorous benchmark for evaluating the reasoning and problem-solving capabilities of large language models

Seedance 2.0: Advancing Video Generation for World Complexity

Model ReleasesDGX agent

arXiv:2604.14148v1 Announce Type: new Abstract: Seedance 2.0 is a new native multi-modal audio-video generation model, officially released in China in early February 2026. Compared with its predecesso

Seek-and-Solve: Benchmarking MLLMs for Visual Clue-Driven Reasoning in Daily Scenarios

Model ReleasesDGX agent

arXiv:2604.14041v1 Announce Type: new Abstract: Daily scenarios are characterized by visual richness, requiring Multimodal Large Language Models (MLLMs) to filter noise and identify decisive visual cl

Self-Organizing Maps with Optimized Latent Positions

Model ReleasesDGX agent

arXiv:2604.13622v1 Announce Type: new Abstract: Self-Organizing Maps (SOM) are a classical method for unsupervised learning, vector quantization, and topographic mapping of high-dimensional data. Howe

SemAttNet: Towards Attention-based Semantic Aware Guided Depth Completion

Model ReleasesDGX agent

arXiv:2204.13635v2 Announce Type: replace Abstract: Depth completion involves recovering a dense depth map from a sparse map and an RGB image. Recent approaches focus on utilizing color images as guid

SHARe-KAN: Post-Training Vector Quantization for Cache-Resident KAN Inference

Model ReleasesDGX agent

arXiv:2512.15742v2 Announce Type: replace Abstract: Pre-trained Vision Kolmogorov-Arnold Networks (KANs) store a dense B-spline grid on every edge, inflating prediction-head parameter counts by more t

Shocking result on my pelican benchmark this morning, I got a better pelican from a 21GB local Qwen3.6-35B-A3B running on my laptop than I d…

Model ReleasesDGX agent

Shocking result on my pelican benchmark this morning, I got a better pelican from a 21GB local Qwen3.6-35B-A3B running on my laptop than I did from the new Opus 4.7! Qwen on the left, Opus on the righ

SLQ: Bridging Modalities via Shared Latent Queries for Retrieval with Frozen MLLMs

Model ReleasesDGX agent

arXiv:2604.13710v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) exhibit strong reasoning and world knowledge, yet adapting them for retrieval remains challenging. Existing app

SocialMirror: Reconstructing 3D Human Interaction Behaviors from Monocular Videos with Semantic and Geometric Guidance

Model ReleasesDGX agent

arXiv:2604.13581v1 Announce Type: new Abstract: Accurately reconstructing human behavior in close-interaction scenarios is crucial for enabling realistic virtual interactions in augmented reality, pre

Sorry for the long wait, everyone! As I said, Qwen is going to keep open-sourcing!

Model ReleasesDGX agent

Sorry for the long wait, everyone! As I said, Qwen is going to keep open-sourcing! ⚡ Meet Qwen3.6-35B-A3B:Now Open-Source!🚀🚀 A sparse MoE model, 35B total params, 3B active. Apache 2.0 license. 🔥 Agen

SparseBalance: Load-Balanced Long Context Training with Dynamic Sparse Attention

Model ReleasesDGX agent

arXiv:2604.13847v1 Announce Type: new Abstract: While sparse attention mitigates the computational bottleneck of long-context LLM training, its distributed training process exhibits extreme heterogene

SpatialEvo: Self-Evolving Spatial Intelligence via Deterministic Geometric Environments

Model ReleasesDGX agent

arXiv:2604.14144v1 Announce Type: cross Abstract: Spatial reasoning over three-dimensional scenes is a core capability for embodied intelligence, yet continuous model improvement remains bottlenecked

Spectral Entropy Collapse as an Empirical Signature of Delayed Generalisation in Grokking

Model ReleasesDGX agent

arXiv:2604.13123v1 Announce Type: new Abstract: Grokking -- delayed generalisation long after memorisation -- lacks a predictive mechanistic explanation. We identify the normalised spectral entropy il

Stein Variational Uncertainty-Adaptive Model Predictive Control

Model ReleasesDGX agent

arXiv:2604.01034v2 Announce Type: replace Abstract: We propose a Stein variational distributionally robust controller for nonlinear dynamical systems with latent parametric uncertainty. The method is

Stochastic Trust-Region Methods for Over-parameterized Models

Model ReleasesDGX agent

arXiv:2604.14017v1 Announce Type: cross Abstract: Under interpolation-type assumptions such as the strong growth condition, stochastic optimization methods can attain convergence rates comparable to f

Syn-TurnTurk: A Synthetic Dataset for Turn-Taking Prediction in Turkish Dialogues

Model ReleasesDGX agent

arXiv:2604.13620v1 Announce Type: new Abstract: Managing natural dialogue timing is a significant challenge for voice-based chatbots. Most current systems usually rely on simple silence detection, whi

Synthesis and Deployment of Maximal Robust Control Barrier Functions through Adversarial Reinforcement Learning

Model ReleasesDGX agent

arXiv:2604.13192v1 Announce Type: cross Abstract: Robust control barrier functions (CBFs) provide a principled mechanism for smooth safety enforcement under worst-case disturbances. However, existing

Synthesizing Instruction-Tuning Datasets with Contrastive Decoding

Model ReleasesDGX agent

arXiv:2604.13538v1 Announce Type: new Abstract: Using responses generated by high-performing large language models (LLMs) for instruction tuning has become a widely adopted approach. However, the exis

Synthetic Tabular Generators Fail to Preserve Behavioral Fraud Patterns: A Benchmark on Temporal, Velocity, and Multi-Account Signals

Model ReleasesDGX agent

arXiv:2604.13125v1 Announce Type: new Abstract: We introduce behavioral fidelity -- a third evaluation dimension for synthetic tabular data that measures whether generated data preserves the temporal,

Target-Bench: Can Video World Models Achieve Mapless Path Planning with Semantic Targets?

Model ReleasesDGX agent

arXiv:2511.17792v2 Announce Type: replace Abstract: While recent video world models can generate highly realistic videos, their ability to perform semantic reasoning and planning remains unclear and u

Text-as-Signal: Quantitative Semantic Scoring with Embeddings, Logprobs, and Noise Reduction

Model ReleasesDGX agent

arXiv:2604.13056v1 Announce Type: new Abstract: This paper presents a practical pipeline for turning text corpora into quantitative semantic signals. Each news item is represented as a full-document e

Text-Attributed Knowledge Graph Enrichment with Large Language Models for Medical Concept Representation

Model ReleasesDGX agent

arXiv:2604.13331v1 Announce Type: new Abstract: In electronic health record (EHR) mining, learning high-quality representations of medical concepts (e.g., standardized diagnosis, medication, and proce

The cognitive companion: a lightweight parallel monitoring architecture for detecting and recovering from reasoning degradation in LLM agents

Model ReleasesDGX agent

arXiv:2604.13759v1 Announce Type: cross Abstract: Large language model (LLM) agents on multi-step tasks suffer reasoning degradation, looping, drift, stuck states, at rates up to 30% on hard tasks. Cu

The Consciousness Cluster: Emergent preferences of Models that Claim to be Conscious

Model ReleasesDGX agent

arXiv:2604.13051v1 Announce Type: new Abstract: There is debate about whether LLMs can be conscious. We investigate a distinct question: if a model claims to be conscious, how does this affect its dow

The UK unveils Sovereign AI, a £500M fund to invest in domestic AI startups, starting with Callosum, which builds software to help different chips work together (Joel Khalili/Wired)

Model ReleasesDGX agent

Joel Khalili / Wired: The UK unveils Sovereign AI, a £500M fund to invest in domestic AI startups, starting with Callosum, which builds software to help different chips work together — In a bid to min

These updates are rolling out on the Codex desktop app starting today. https://openai.com/index/codex-for-almost-everything/

Model ReleasesDGX agent

OpenAI announced the rollout of updates to the Codex desktop application, beginning on the date of the announcement. The updates likely enhance Codex's code generation and AI-assisted programming capa

TIP: Token Importance in On-Policy Distillation

Model ReleasesDGX agent

arXiv:2604.14084v1 Announce Type: new Abstract: On-policy knowledge distillation (OPD) trains a student on its own rollouts under token-level supervision from a teacher. Not all token positions matter

TLoRA+: A Low-Rank Parameter-Efficient Fine-Tuning Method for Large Language Models

Model ReleasesDGX agent

arXiv:2604.13368v1 Announce Type: new Abstract: Fine-tuning large language models (LLMs) aims to adapt pre-trained models to specific tasks using relatively small and domain-specific datasets. Among P

Today we’re announcing Ternary Bonsai: Top intelligence at 1.58 bits Using ternary weights {-1, 0, +1}, we built a family of models that are…

Model ReleasesDGX agent

Today we’re announcing Ternary Bonsai: Top intelligence at 1.58 bits Using ternary weights {-1, 0, +1}, we built a family of models that are 9x smaller than their 16-bit counterparts while outperformi

ToolOmni: Enabling Open-World Tool Use via Agentic learning with Proactive Retrieval and Grounded Execution

Model ReleasesDGX agent

arXiv:2604.13787v1 Announce Type: new Abstract: Large Language Models (LLMs) enhance their problem-solving capability by utilizing external tools. However, in open-world scenarios with massive and evo

Towards Generalizable Robotic Manipulation in Dynamic Environments

Model ReleasesDGX agent

arXiv:2603.15620v2 Announce Type: replace Abstract: Vision-Language-Action (VLA) models excel in static manipulation but struggle in dynamic environments with moving targets. This performance gap prim

Towards Successful Implementation of Automated Raveling Detection: Effects of Training Data Size, Illumination Difference, and Spatial Shift

Model ReleasesDGX agent

arXiv:2604.13322v1 Announce Type: new Abstract: Raveling, the loss of aggregates, is a major form of asphalt pavement surface distress, especially on highways. While research has shown that machine le

Transcriptomic Models for Immunotherapy Response Prediction Show Limited Cross-cohort Generalisability

Model ReleasesDGX agent

arXiv:2604.05478v2 Announce Type: replace-cross Abstract: Immune checkpoint inhibitors (ICIs) have transformed cancer therapy; yet substantial proportion of patients exhibit intrinsic or acquired resi

Treating enterprise AI as an operating layer

Model ReleasesDGX agent

There’s a fault line running through enterprise AI, and it’s not the one getting the most attention. The public conversation still tracks foundation models and benchmarks—GPT versus Gemini, reasoning

TREX: Automating LLM Fine-tuning via Agent-Driven Tree-based Exploration

Model ReleasesDGX agent

arXiv:2604.14116v1 Announce Type: cross Abstract: While Large Language Models (LLMs) have empowered AI research agents to perform isolated scientific tasks, automating complex, real-world workflows, s

Two-Stage Regularization-Based Structured Pruning for LLMs

Model ReleasesDGX agent

arXiv:2505.18232v3 Announce Type: replace-cross Abstract: The deployment of large language models (LLMs) is largely hindered by their large number of parameters. Structural pruning has emerged as a pr

UI-Copilot: Advancing Long-Horizon GUI Automation via Tool-Integrated Policy Optimization

Model ReleasesDGX agent

arXiv:2604.13822v1 Announce Type: new Abstract: MLLM-based GUI agents have demonstrated strong capabilities in complex user interface interaction tasks. However, long-horizon scenarios remain challeng

UniBlendNet: Unified Global, Multi-Scale, and Region-Adaptive Modeling for Ambient Lighting Normalization

Model ReleasesDGX agent

arXiv:2604.13383v1 Announce Type: new Abstract: Ambient Lighting Normalization (ALN) aims to restore images degraded by complex, spatially varying illumination conditions. Existing methods, such as IF

UniGeoSeg: Towards Unified Open-World Segmentation for Geospatial Scenes

Model ReleasesDGX agent

arXiv:2511.23332v2 Announce Type: replace Abstract: Instruction-driven segmentation in remote sensing generates masks from guidance, offering great potential for accessible and generalizable applicati

UNRIO: Uncertainty-Aware Velocity Learning for Radar-Inertial Odometry

Model ReleasesDGX agent

arXiv:2604.13584v1 Announce Type: new Abstract: We present UNRIO, an uncertainty-aware radar-inertial odometry system that estimates ego-velocity directly from raw mmWave radar IQ signals rather than

Unsupervised Anomaly Detection in Process-Complex Industrial Time Series: A Real-World Case Study

Model ReleasesDGX agent

arXiv:2604.13928v1 Announce Type: new Abstract: Industrial time-series data from real production environments exhibits substantially higher complexity than commonly used benchmark datasets, primarily

ValueGround: Evaluating Culture-Conditioned Visual Value Grounding in MLLMs

Model ReleasesDGX agent

arXiv:2604.06484v2 Announce Type: replace Abstract: Cultural values are expressed not only through language but also through visual scenes and everyday social practices. Yet existing evaluations of cu

VGGT-Segmentor: Geometry-Enhanced Cross-View Segmentation

Model ReleasesDGX agent

arXiv:2604.13596v1 Announce Type: new Abstract: Instance-level object segmentation across disparate egocentric and exocentric views is a fundamental challenge in visual understanding, critical for app

ViBES: A Conversational Agent with Behaviorally-Intelligent 3D Virtual Body

Model ReleasesDGX agent

arXiv:2512.14234v2 Announce Type: replace Abstract: Human communication is inherently multimodal and social: words, prosody, and body language jointly carry intent. Yet most prior systems model human

VLM Performance:Qwen3.6 is natively multimodal, and Qwen3.6-35B-A3B showcases perception and multimodal reasoning capabilities that far exce…

Model ReleasesDGX agent

VLM Performance:Qwen3.6 is natively multimodal, and Qwen3.6-35B-A3B showcases perception and multimodal reasoning capabilities that far exceed what its size would suggest, with only around 3 billion a

WAI-ANIMA 1.0 released

Model ReleasesDGX agent

WAI-ANIMA 1.0 is a newly released Stable Diffusion checkpoint model from the WAI model family, likely combining elements of the WAI-Illustrious anime generation lineage with the Anima diffusion archit

We comprehensively benchmarked Opus 4.7 on document understanding. We evaluated it through ParseBench - our comprehensive OCR benchmark for …

Model ReleasesDGX agent

We comprehensively benchmarked Opus 4.7 on document understanding. We evaluated it through ParseBench - our comprehensive OCR benchmark for enterprise documents where we evaluate tables, text, charts,

We fixed a bug where rate limits on Claude subscriptions weren't properly adjusted for long context requests in Opus 4.7. We've reset 5-hour…

Model ReleasesDGX agent

Anthropic fixed a bug in Claude Opus 4.7 where rate limits for paid subscriptions weren't correctly adjusted for requests using the model's extended context window capabilities. The fix involved reset

We replicated Mythos findings in opencode using public models, not Anthropic's private stack. The moat is moving from model access to valida…

Model ReleasesDGX agent

We replicated Mythos findings in opencode using public models, not Anthropic's private stack. The moat is moving from model access to validation: finding vulnerability signal is getting cheaper; turni

'We're not looking for the easiest path, we're looking for the right path.' NEW: Aidan Gomez (@aidangomez) is trying to build a very differe…

Model ReleasesDGX agent

'We're not looking for the easiest path, we're looking for the right path.' NEW: Aidan Gomez (@aidangomez) is trying to build a very different kind of AI company at Cohere. Based in Toronto, $6.8B-val

We’re running monthly ‘what we shipped’ webinars. You can sign up for that and our other our webinars here: https://www.anthropic.com/webina…

Model ReleasesDGX agent

Anthropic hosts monthly webinars where they present product updates and new features they've released, with signup information available on their webinars page. The company offers multiple webinar ser

We’ve also added support for 90+ plugins in Codex, giving it more ways to gather context and take action across the tools you already use fo…

Model ReleasesDGX agent

We’ve also added support for 90+ plugins in Codex, giving it more ways to gather context and take action across the tools you already use for docs, project management, code review, creative work, depl

We’ve heard your feedback and we’re working on making it easier to follow everything that’s happening with Claude Code. First, we’re introdu…

Model ReleasesDGX agent

We’ve heard your feedback and we’re working on making it easier to follow everything that’s happening with Claude Code. First, we’re introducing @ClaudeDevs, the official channel to follow for all upd

What you need to know about Opus 4.7 * Takes instructions literally * Better vision means improved computer use and producing slides and oth…

Model ReleasesDGX agent

What you need to know about Opus 4.7 * Takes instructions literally * Better vision means improved computer use and producing slides and other visual artifacts * Optimized for large-scale real-world a

What’s New in Microsoft Foundry Fine-Tuning | April 2026

Model ReleasesDGX agent

April 2026 brings three major Reinforcement Fine-Tuning updates: Global Training for o4-mini with lower per-token rates across 12+ regions, new GPT-4.1 model graders for richer reward signals, and a c

← Previous
1…342343344345346…372
Next →