AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
All
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,292 results
Model Releases

RPS: Information Elicitation with Reinforcement Prompt Selection

DGX agent

arXiv:2604.13817v1 Announce Type: new Abstract: Large language models (LLMs) have shown remarkable capabilities in dialogue generation and reasoning, yet their effectiveness in eliciting user-known bu

model-releasesarxiv-cs-lg
16 Apr 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Scaling Test-Time Compute to Achieve IOI Gold Medal with Open-Weight Models

DGX agent

arXiv:2510.14232v2 Announce Type: replace-cross Abstract: Competitive programming has become a rigorous benchmark for evaluating the reasoning and problem-solving capabilities of large language models

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Seedance 2.0: Advancing Video Generation for World Complexity

DGX agent

arXiv:2604.14148v1 Announce Type: new Abstract: Seedance 2.0 is a new native multi-modal audio-video generation model, officially released in China in early February 2026. Compared with its predecesso

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

Seek-and-Solve: Benchmarking MLLMs for Visual Clue-Driven Reasoning in Daily Scenarios

DGX agent

arXiv:2604.14041v1 Announce Type: new Abstract: Daily scenarios are characterized by visual richness, requiring Multimodal Large Language Models (MLLMs) to filter noise and identify decisive visual cl

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

Self-Organizing Maps with Optimized Latent Positions

DGX agent

arXiv:2604.13622v1 Announce Type: new Abstract: Self-Organizing Maps (SOM) are a classical method for unsupervised learning, vector quantization, and topographic mapping of high-dimensional data. Howe

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

SemAttNet: Towards Attention-based Semantic Aware Guided Depth Completion

DGX agent

arXiv:2204.13635v2 Announce Type: replace Abstract: Depth completion involves recovering a dense depth map from a sparse map and an RGB image. Recent approaches focus on utilizing color images as guid

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

SHARe-KAN: Post-Training Vector Quantization for Cache-Resident KAN Inference

DGX agent

arXiv:2512.15742v2 Announce Type: replace Abstract: Pre-trained Vision Kolmogorov-Arnold Networks (KANs) store a dense B-spline grid on every edge, inflating prediction-head parameter counts by more t

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

Shocking result on my pelican benchmark this morning, I got a better pelican from a 21GB local Qwen3.6-35B-A3B running on my laptop than I d…

DGX agent

Shocking result on my pelican benchmark this morning, I got a better pelican from a 21GB local Qwen3.6-35B-A3B running on my laptop than I did from the new Opus 4.7! Qwen on the left, Opus on the righ

model-releasesclem-delangue--x
16 Apr 2026
Model Releases

SLQ: Bridging Modalities via Shared Latent Queries for Retrieval with Frozen MLLMs

DGX agent

arXiv:2604.13710v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) exhibit strong reasoning and world knowledge, yet adapting them for retrieval remains challenging. Existing app

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

SocialMirror: Reconstructing 3D Human Interaction Behaviors from Monocular Videos with Semantic and Geometric Guidance

DGX agent

arXiv:2604.13581v1 Announce Type: new Abstract: Accurately reconstructing human behavior in close-interaction scenarios is crucial for enabling realistic virtual interactions in augmented reality, pre

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

Sorry for the long wait, everyone! As I said, Qwen is going to keep open-sourcing!

DGX agent

Sorry for the long wait, everyone! As I said, Qwen is going to keep open-sourcing! ⚡ Meet Qwen3.6-35B-A3B:Now Open-Source!🚀🚀 A sparse MoE model, 35B total params, 3B active. Apache 2.0 license. 🔥 Agen

model-releasesclem-delangue--x
16 Apr 2026
Model Releases

SparseBalance: Load-Balanced Long Context Training with Dynamic Sparse Attention

DGX agent

arXiv:2604.13847v1 Announce Type: new Abstract: While sparse attention mitigates the computational bottleneck of long-context LLM training, its distributed training process exhibits extreme heterogene

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

SpatialEvo: Self-Evolving Spatial Intelligence via Deterministic Geometric Environments

DGX agent

arXiv:2604.14144v1 Announce Type: cross Abstract: Spatial reasoning over three-dimensional scenes is a core capability for embodied intelligence, yet continuous model improvement remains bottlenecked

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Spectral Entropy Collapse as an Empirical Signature of Delayed Generalisation in Grokking

DGX agent

arXiv:2604.13123v1 Announce Type: new Abstract: Grokking -- delayed generalisation long after memorisation -- lacks a predictive mechanistic explanation. We identify the normalised spectral entropy il

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

Stein Variational Uncertainty-Adaptive Model Predictive Control

DGX agent

arXiv:2604.01034v2 Announce Type: replace Abstract: We propose a Stein variational distributionally robust controller for nonlinear dynamical systems with latent parametric uncertainty. The method is

model-releasesarxiv-cs-ro
16 Apr 2026
Model Releases

Stochastic Trust-Region Methods for Over-parameterized Models

DGX agent

arXiv:2604.14017v1 Announce Type: cross Abstract: Under interpolation-type assumptions such as the strong growth condition, stochastic optimization methods can attain convergence rates comparable to f

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

Syn-TurnTurk: A Synthetic Dataset for Turn-Taking Prediction in Turkish Dialogues

DGX agent

arXiv:2604.13620v1 Announce Type: new Abstract: Managing natural dialogue timing is a significant challenge for voice-based chatbots. Most current systems usually rely on simple silence detection, whi

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Synthesis and Deployment of Maximal Robust Control Barrier Functions through Adversarial Reinforcement Learning

DGX agent

arXiv:2604.13192v1 Announce Type: cross Abstract: Robust control barrier functions (CBFs) provide a principled mechanism for smooth safety enforcement under worst-case disturbances. However, existing

model-releasesarxiv-cs-ro
16 Apr 2026
Model Releases

Synthesizing Instruction-Tuning Datasets with Contrastive Decoding

DGX agent

arXiv:2604.13538v1 Announce Type: new Abstract: Using responses generated by high-performing large language models (LLMs) for instruction tuning has become a widely adopted approach. However, the exis

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Synthetic Tabular Generators Fail to Preserve Behavioral Fraud Patterns: A Benchmark on Temporal, Velocity, and Multi-Account Signals

DGX agent

arXiv:2604.13125v1 Announce Type: new Abstract: We introduce behavioral fidelity -- a third evaluation dimension for synthetic tabular data that measures whether generated data preserves the temporal,

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

Target-Bench: Can Video World Models Achieve Mapless Path Planning with Semantic Targets?

DGX agent

arXiv:2511.17792v2 Announce Type: replace Abstract: While recent video world models can generate highly realistic videos, their ability to perform semantic reasoning and planning remains unclear and u

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

Text-as-Signal: Quantitative Semantic Scoring with Embeddings, Logprobs, and Noise Reduction

DGX agent

arXiv:2604.13056v1 Announce Type: new Abstract: This paper presents a practical pipeline for turning text corpora into quantitative semantic signals. Each news item is represented as a full-document e

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Text-Attributed Knowledge Graph Enrichment with Large Language Models for Medical Concept Representation

DGX agent

arXiv:2604.13331v1 Announce Type: new Abstract: In electronic health record (EHR) mining, learning high-quality representations of medical concepts (e.g., standardized diagnosis, medication, and proce

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

The cognitive companion: a lightweight parallel monitoring architecture for detecting and recovering from reasoning degradation in LLM agents

DGX agent

arXiv:2604.13759v1 Announce Type: cross Abstract: Large language model (LLM) agents on multi-step tasks suffer reasoning degradation, looping, drift, stuck states, at rates up to 30% on hard tasks. Cu

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

The Consciousness Cluster: Emergent preferences of Models that Claim to be Conscious

DGX agent

arXiv:2604.13051v1 Announce Type: new Abstract: There is debate about whether LLMs can be conscious. We investigate a distinct question: if a model claims to be conscious, how does this affect its dow

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

The UK unveils Sovereign AI, a £500M fund to invest in domestic AI startups, starting with Callosum, which builds software to help different chips work together (Joel Khalili/Wired)

DGX agent

Joel Khalili / Wired: The UK unveils Sovereign AI, a £500M fund to invest in domestic AI startups, starting with Callosum, which builds software to help different chips work together — In a bid to min

model-releasestechmeme
16 Apr 2026
Model Releases

These updates are rolling out on the Codex desktop app starting today. https://openai.com/index/codex-for-almost-everything/

DGX agent

OpenAI announced the rollout of updates to the Codex desktop application, beginning on the date of the announcement. The updates likely enhance Codex's code generation and AI-assisted programming capa

model-releasesopenai--x
16 Apr 2026
Model Releases

TIP: Token Importance in On-Policy Distillation

DGX agent

arXiv:2604.14084v1 Announce Type: new Abstract: On-policy knowledge distillation (OPD) trains a student on its own rollouts under token-level supervision from a teacher. Not all token positions matter

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

TLoRA+: A Low-Rank Parameter-Efficient Fine-Tuning Method for Large Language Models

DGX agent

arXiv:2604.13368v1 Announce Type: new Abstract: Fine-tuning large language models (LLMs) aims to adapt pre-trained models to specific tasks using relatively small and domain-specific datasets. Among P

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Today we’re announcing Ternary Bonsai: Top intelligence at 1.58 bits Using ternary weights {-1, 0, +1}, we built a family of models that are…

DGX agent

Today we’re announcing Ternary Bonsai: Top intelligence at 1.58 bits Using ternary weights {-1, 0, +1}, we built a family of models that are 9x smaller than their 16-bit counterparts while outperformi

model-releasesemad-mostaque--x
16 Apr 2026
Model Releases

ToolOmni: Enabling Open-World Tool Use via Agentic learning with Proactive Retrieval and Grounded Execution

DGX agent

arXiv:2604.13787v1 Announce Type: new Abstract: Large Language Models (LLMs) enhance their problem-solving capability by utilizing external tools. However, in open-world scenarios with massive and evo

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Towards Generalizable Robotic Manipulation in Dynamic Environments

DGX agent

arXiv:2603.15620v2 Announce Type: replace Abstract: Vision-Language-Action (VLA) models excel in static manipulation but struggle in dynamic environments with moving targets. This performance gap prim

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

Towards Successful Implementation of Automated Raveling Detection: Effects of Training Data Size, Illumination Difference, and Spatial Shift

DGX agent

arXiv:2604.13322v1 Announce Type: new Abstract: Raveling, the loss of aggregates, is a major form of asphalt pavement surface distress, especially on highways. While research has shown that machine le

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

Transcriptomic Models for Immunotherapy Response Prediction Show Limited Cross-cohort Generalisability

DGX agent

arXiv:2604.05478v2 Announce Type: replace-cross Abstract: Immune checkpoint inhibitors (ICIs) have transformed cancer therapy; yet substantial proportion of patients exhibit intrinsic or acquired resi

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

Treating enterprise AI as an operating layer

DGX agent

There’s a fault line running through enterprise AI, and it’s not the one getting the most attention. The public conversation still tracks foundation models and benchmarks—GPT versus Gemini, reasoning

model-releasesmit-tech-review
16 Apr 2026
Model Releases

TREX: Automating LLM Fine-tuning via Agent-Driven Tree-based Exploration

DGX agent

arXiv:2604.14116v1 Announce Type: cross Abstract: While Large Language Models (LLMs) have empowered AI research agents to perform isolated scientific tasks, automating complex, real-world workflows, s

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Two-Stage Regularization-Based Structured Pruning for LLMs

DGX agent

arXiv:2505.18232v3 Announce Type: replace-cross Abstract: The deployment of large language models (LLMs) is largely hindered by their large number of parameters. Structural pruning has emerged as a pr

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

UI-Copilot: Advancing Long-Horizon GUI Automation via Tool-Integrated Policy Optimization

DGX agent

arXiv:2604.13822v1 Announce Type: new Abstract: MLLM-based GUI agents have demonstrated strong capabilities in complex user interface interaction tasks. However, long-horizon scenarios remain challeng

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

UniBlendNet: Unified Global, Multi-Scale, and Region-Adaptive Modeling for Ambient Lighting Normalization

DGX agent

arXiv:2604.13383v1 Announce Type: new Abstract: Ambient Lighting Normalization (ALN) aims to restore images degraded by complex, spatially varying illumination conditions. Existing methods, such as IF

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

UniGeoSeg: Towards Unified Open-World Segmentation for Geospatial Scenes

DGX agent

arXiv:2511.23332v2 Announce Type: replace Abstract: Instruction-driven segmentation in remote sensing generates masks from guidance, offering great potential for accessible and generalizable applicati

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

UNRIO: Uncertainty-Aware Velocity Learning for Radar-Inertial Odometry

DGX agent

arXiv:2604.13584v1 Announce Type: new Abstract: We present UNRIO, an uncertainty-aware radar-inertial odometry system that estimates ego-velocity directly from raw mmWave radar IQ signals rather than

model-releasesarxiv-cs-ro
16 Apr 2026
Model Releases

Unsupervised Anomaly Detection in Process-Complex Industrial Time Series: A Real-World Case Study

DGX agent

arXiv:2604.13928v1 Announce Type: new Abstract: Industrial time-series data from real production environments exhibits substantially higher complexity than commonly used benchmark datasets, primarily

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

ValueGround: Evaluating Culture-Conditioned Visual Value Grounding in MLLMs

DGX agent

arXiv:2604.06484v2 Announce Type: replace Abstract: Cultural values are expressed not only through language but also through visual scenes and everyday social practices. Yet existing evaluations of cu

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

VGGT-Segmentor: Geometry-Enhanced Cross-View Segmentation

DGX agent

arXiv:2604.13596v1 Announce Type: new Abstract: Instance-level object segmentation across disparate egocentric and exocentric views is a fundamental challenge in visual understanding, critical for app

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

ViBES: A Conversational Agent with Behaviorally-Intelligent 3D Virtual Body

DGX agent

arXiv:2512.14234v2 Announce Type: replace Abstract: Human communication is inherently multimodal and social: words, prosody, and body language jointly carry intent. Yet most prior systems model human

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

VLM Performance:Qwen3.6 is natively multimodal, and Qwen3.6-35B-A3B showcases perception and multimodal reasoning capabilities that far exce…

DGX agent

VLM Performance:Qwen3.6 is natively multimodal, and Qwen3.6-35B-A3B showcases perception and multimodal reasoning capabilities that far exceed what its size would suggest, with only around 3 billion a

model-releasesqwen--x
16 Apr 2026
Model Releases

WAI-ANIMA 1.0 released

DGX agent

WAI-ANIMA 1.0 is a newly released Stable Diffusion checkpoint model from the WAI model family, likely combining elements of the WAI-Illustrious anime generation lineage with the Anima diffusion archit

model-releasesr-stablediffusion
16 Apr 2026
Model Releases

We comprehensively benchmarked Opus 4.7 on document understanding. We evaluated it through ParseBench - our comprehensive OCR benchmark for …

DGX agent

We comprehensively benchmarked Opus 4.7 on document understanding. We evaluated it through ParseBench - our comprehensive OCR benchmark for enterprise documents where we evaluate tables, text, charts,

model-releasesjerry-liu--x
16 Apr 2026
← Previous
1…429430431432433…465
Next →