AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,284 results
Model Releases

SLQ: Bridging Modalities via Shared Latent Queries for Retrieval with Frozen MLLMs

DGX agent

arXiv:2604.13710v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) exhibit strong reasoning and world knowledge, yet adapting them for retrieval remains challenging. Existing app

model-releasesarxiv-cs-cv
16 Apr 2026
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

SocialMirror: Reconstructing 3D Human Interaction Behaviors from Monocular Videos with Semantic and Geometric Guidance

DGX agent

arXiv:2604.13581v1 Announce Type: new Abstract: Accurately reconstructing human behavior in close-interaction scenarios is crucial for enabling realistic virtual interactions in augmented reality, pre

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

Sorry for the long wait, everyone! As I said, Qwen is going to keep open-sourcing!

DGX agent

Sorry for the long wait, everyone! As I said, Qwen is going to keep open-sourcing! ⚡ Meet Qwen3.6-35B-A3B:Now Open-Source!🚀🚀 A sparse MoE model, 35B total params, 3B active. Apache 2.0 license. 🔥 Agen

model-releasesclem-delangue--x
16 Apr 2026
Model Releases

SparseBalance: Load-Balanced Long Context Training with Dynamic Sparse Attention

DGX agent

arXiv:2604.13847v1 Announce Type: new Abstract: While sparse attention mitigates the computational bottleneck of long-context LLM training, its distributed training process exhibits extreme heterogene

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

SpatialEvo: Self-Evolving Spatial Intelligence via Deterministic Geometric Environments

DGX agent

arXiv:2604.14144v1 Announce Type: cross Abstract: Spatial reasoning over three-dimensional scenes is a core capability for embodied intelligence, yet continuous model improvement remains bottlenecked

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Spectral Entropy Collapse as an Empirical Signature of Delayed Generalisation in Grokking

DGX agent

arXiv:2604.13123v1 Announce Type: new Abstract: Grokking -- delayed generalisation long after memorisation -- lacks a predictive mechanistic explanation. We identify the normalised spectral entropy il

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

Stein Variational Uncertainty-Adaptive Model Predictive Control

DGX agent

arXiv:2604.01034v2 Announce Type: replace Abstract: We propose a Stein variational distributionally robust controller for nonlinear dynamical systems with latent parametric uncertainty. The method is

model-releasesarxiv-cs-ro
16 Apr 2026
Model Releases

Stochastic Trust-Region Methods for Over-parameterized Models

DGX agent

arXiv:2604.14017v1 Announce Type: cross Abstract: Under interpolation-type assumptions such as the strong growth condition, stochastic optimization methods can attain convergence rates comparable to f

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

Syn-TurnTurk: A Synthetic Dataset for Turn-Taking Prediction in Turkish Dialogues

DGX agent

arXiv:2604.13620v1 Announce Type: new Abstract: Managing natural dialogue timing is a significant challenge for voice-based chatbots. Most current systems usually rely on simple silence detection, whi

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Synthesis and Deployment of Maximal Robust Control Barrier Functions through Adversarial Reinforcement Learning

DGX agent

arXiv:2604.13192v1 Announce Type: cross Abstract: Robust control barrier functions (CBFs) provide a principled mechanism for smooth safety enforcement under worst-case disturbances. However, existing

model-releasesarxiv-cs-ro
16 Apr 2026
Model Releases

Synthesizing Instruction-Tuning Datasets with Contrastive Decoding

DGX agent

arXiv:2604.13538v1 Announce Type: new Abstract: Using responses generated by high-performing large language models (LLMs) for instruction tuning has become a widely adopted approach. However, the exis

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Synthetic Tabular Generators Fail to Preserve Behavioral Fraud Patterns: A Benchmark on Temporal, Velocity, and Multi-Account Signals

DGX agent

arXiv:2604.13125v1 Announce Type: new Abstract: We introduce behavioral fidelity -- a third evaluation dimension for synthetic tabular data that measures whether generated data preserves the temporal,

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

Target-Bench: Can Video World Models Achieve Mapless Path Planning with Semantic Targets?

DGX agent

arXiv:2511.17792v2 Announce Type: replace Abstract: While recent video world models can generate highly realistic videos, their ability to perform semantic reasoning and planning remains unclear and u

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

Text-as-Signal: Quantitative Semantic Scoring with Embeddings, Logprobs, and Noise Reduction

DGX agent

arXiv:2604.13056v1 Announce Type: new Abstract: This paper presents a practical pipeline for turning text corpora into quantitative semantic signals. Each news item is represented as a full-document e

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Text-Attributed Knowledge Graph Enrichment with Large Language Models for Medical Concept Representation

DGX agent

arXiv:2604.13331v1 Announce Type: new Abstract: In electronic health record (EHR) mining, learning high-quality representations of medical concepts (e.g., standardized diagnosis, medication, and proce

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

The cognitive companion: a lightweight parallel monitoring architecture for detecting and recovering from reasoning degradation in LLM agents

DGX agent

arXiv:2604.13759v1 Announce Type: cross Abstract: Large language model (LLM) agents on multi-step tasks suffer reasoning degradation, looping, drift, stuck states, at rates up to 30% on hard tasks. Cu

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

The Consciousness Cluster: Emergent preferences of Models that Claim to be Conscious

DGX agent

arXiv:2604.13051v1 Announce Type: new Abstract: There is debate about whether LLMs can be conscious. We investigate a distinct question: if a model claims to be conscious, how does this affect its dow

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

The UK unveils Sovereign AI, a £500M fund to invest in domestic AI startups, starting with Callosum, which builds software to help different chips work together (Joel Khalili/Wired)

DGX agent

Joel Khalili / Wired: The UK unveils Sovereign AI, a £500M fund to invest in domestic AI startups, starting with Callosum, which builds software to help different chips work together — In a bid to min

model-releasestechmeme
16 Apr 2026
Model Releases

These updates are rolling out on the Codex desktop app starting today. https://openai.com/index/codex-for-almost-everything/

DGX agent

OpenAI announced the rollout of updates to the Codex desktop application, beginning on the date of the announcement. The updates likely enhance Codex's code generation and AI-assisted programming capa

model-releasesopenai--x
16 Apr 2026
Model Releases

TIP: Token Importance in On-Policy Distillation

DGX agent

arXiv:2604.14084v1 Announce Type: new Abstract: On-policy knowledge distillation (OPD) trains a student on its own rollouts under token-level supervision from a teacher. Not all token positions matter

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

TLoRA+: A Low-Rank Parameter-Efficient Fine-Tuning Method for Large Language Models

DGX agent

arXiv:2604.13368v1 Announce Type: new Abstract: Fine-tuning large language models (LLMs) aims to adapt pre-trained models to specific tasks using relatively small and domain-specific datasets. Among P

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Today we’re announcing Ternary Bonsai: Top intelligence at 1.58 bits Using ternary weights {-1, 0, +1}, we built a family of models that are…

DGX agent

Today we’re announcing Ternary Bonsai: Top intelligence at 1.58 bits Using ternary weights {-1, 0, +1}, we built a family of models that are 9x smaller than their 16-bit counterparts while outperformi

model-releasesemad-mostaque--x
16 Apr 2026
Model Releases

ToolOmni: Enabling Open-World Tool Use via Agentic learning with Proactive Retrieval and Grounded Execution

DGX agent

arXiv:2604.13787v1 Announce Type: new Abstract: Large Language Models (LLMs) enhance their problem-solving capability by utilizing external tools. However, in open-world scenarios with massive and evo

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Towards Generalizable Robotic Manipulation in Dynamic Environments

DGX agent

arXiv:2603.15620v2 Announce Type: replace Abstract: Vision-Language-Action (VLA) models excel in static manipulation but struggle in dynamic environments with moving targets. This performance gap prim

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

Towards Successful Implementation of Automated Raveling Detection: Effects of Training Data Size, Illumination Difference, and Spatial Shift

DGX agent

arXiv:2604.13322v1 Announce Type: new Abstract: Raveling, the loss of aggregates, is a major form of asphalt pavement surface distress, especially on highways. While research has shown that machine le

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

Transcriptomic Models for Immunotherapy Response Prediction Show Limited Cross-cohort Generalisability

DGX agent

arXiv:2604.05478v2 Announce Type: replace-cross Abstract: Immune checkpoint inhibitors (ICIs) have transformed cancer therapy; yet substantial proportion of patients exhibit intrinsic or acquired resi

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

Treating enterprise AI as an operating layer

DGX agent

There’s a fault line running through enterprise AI, and it’s not the one getting the most attention. The public conversation still tracks foundation models and benchmarks—GPT versus Gemini, reasoning

model-releasesmit-tech-review
16 Apr 2026
Model Releases

TREX: Automating LLM Fine-tuning via Agent-Driven Tree-based Exploration

DGX agent

arXiv:2604.14116v1 Announce Type: cross Abstract: While Large Language Models (LLMs) have empowered AI research agents to perform isolated scientific tasks, automating complex, real-world workflows, s

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Two-Stage Regularization-Based Structured Pruning for LLMs

DGX agent

arXiv:2505.18232v3 Announce Type: replace-cross Abstract: The deployment of large language models (LLMs) is largely hindered by their large number of parameters. Structural pruning has emerged as a pr

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

UI-Copilot: Advancing Long-Horizon GUI Automation via Tool-Integrated Policy Optimization

DGX agent

arXiv:2604.13822v1 Announce Type: new Abstract: MLLM-based GUI agents have demonstrated strong capabilities in complex user interface interaction tasks. However, long-horizon scenarios remain challeng

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

UniBlendNet: Unified Global, Multi-Scale, and Region-Adaptive Modeling for Ambient Lighting Normalization

DGX agent

arXiv:2604.13383v1 Announce Type: new Abstract: Ambient Lighting Normalization (ALN) aims to restore images degraded by complex, spatially varying illumination conditions. Existing methods, such as IF

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

UniGeoSeg: Towards Unified Open-World Segmentation for Geospatial Scenes

DGX agent

arXiv:2511.23332v2 Announce Type: replace Abstract: Instruction-driven segmentation in remote sensing generates masks from guidance, offering great potential for accessible and generalizable applicati

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

UNRIO: Uncertainty-Aware Velocity Learning for Radar-Inertial Odometry

DGX agent

arXiv:2604.13584v1 Announce Type: new Abstract: We present UNRIO, an uncertainty-aware radar-inertial odometry system that estimates ego-velocity directly from raw mmWave radar IQ signals rather than

model-releasesarxiv-cs-ro
16 Apr 2026
Model Releases

Unsupervised Anomaly Detection in Process-Complex Industrial Time Series: A Real-World Case Study

DGX agent

arXiv:2604.13928v1 Announce Type: new Abstract: Industrial time-series data from real production environments exhibits substantially higher complexity than commonly used benchmark datasets, primarily

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

ValueGround: Evaluating Culture-Conditioned Visual Value Grounding in MLLMs

DGX agent

arXiv:2604.06484v2 Announce Type: replace Abstract: Cultural values are expressed not only through language but also through visual scenes and everyday social practices. Yet existing evaluations of cu

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

VGGT-Segmentor: Geometry-Enhanced Cross-View Segmentation

DGX agent

arXiv:2604.13596v1 Announce Type: new Abstract: Instance-level object segmentation across disparate egocentric and exocentric views is a fundamental challenge in visual understanding, critical for app

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

ViBES: A Conversational Agent with Behaviorally-Intelligent 3D Virtual Body

DGX agent

arXiv:2512.14234v2 Announce Type: replace Abstract: Human communication is inherently multimodal and social: words, prosody, and body language jointly carry intent. Yet most prior systems model human

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

VLM Performance:Qwen3.6 is natively multimodal, and Qwen3.6-35B-A3B showcases perception and multimodal reasoning capabilities that far exce…

DGX agent

VLM Performance:Qwen3.6 is natively multimodal, and Qwen3.6-35B-A3B showcases perception and multimodal reasoning capabilities that far exceed what its size would suggest, with only around 3 billion a

model-releasesqwen--x
16 Apr 2026
Model Releases

WAI-ANIMA 1.0 released

DGX agent

WAI-ANIMA 1.0 is a newly released Stable Diffusion checkpoint model from the WAI model family, likely combining elements of the WAI-Illustrious anime generation lineage with the Anima diffusion archit

model-releasesr-stablediffusion
16 Apr 2026
Model Releases

We comprehensively benchmarked Opus 4.7 on document understanding. We evaluated it through ParseBench - our comprehensive OCR benchmark for …

DGX agent

We comprehensively benchmarked Opus 4.7 on document understanding. We evaluated it through ParseBench - our comprehensive OCR benchmark for enterprise documents where we evaluate tables, text, charts,

model-releasesjerry-liu--x
16 Apr 2026
Model Releases

We fixed a bug where rate limits on Claude subscriptions weren't properly adjusted for long context requests in Opus 4.7. We've reset 5-hour…

DGX agent

Anthropic fixed a bug in Claude Opus 4.7 where rate limits for paid subscriptions weren't correctly adjusted for requests using the model's extended context window capabilities. The fix involved reset

model-releasesboris-cherny--x
16 Apr 2026
Model Releases

We replicated Mythos findings in opencode using public models, not Anthropic's private stack. The moat is moving from model access to valida…

DGX agent

We replicated Mythos findings in opencode using public models, not Anthropic's private stack. The moat is moving from model access to validation: finding vulnerability signal is getting cheaper; turni

model-releasesclem-delangue--x
16 Apr 2026
Model Releases

'We're not looking for the easiest path, we're looking for the right path.' NEW: Aidan Gomez (@aidangomez) is trying to build a very differe…

DGX agent

'We're not looking for the easiest path, we're looking for the right path.' NEW: Aidan Gomez (@aidangomez) is trying to build a very different kind of AI company at Cohere. Based in Toronto, $6.8B-val

model-releasescohere--x
16 Apr 2026
Model Releases

We’re running monthly ‘what we shipped’ webinars. You can sign up for that and our other our webinars here: https://www.anthropic.com/webina…

DGX agent

Anthropic hosts monthly webinars where they present product updates and new features they've released, with signup information available on their webinars page. The company offers multiple webinar ser

model-releasesthariq--x
16 Apr 2026
Model Releases

We’ve also added support for 90+ plugins in Codex, giving it more ways to gather context and take action across the tools you already use fo…

DGX agent

We’ve also added support for 90+ plugins in Codex, giving it more ways to gather context and take action across the tools you already use for docs, project management, code review, creative work, depl

model-releasesopenai--x
16 Apr 2026
Model Releases

We’ve heard your feedback and we’re working on making it easier to follow everything that’s happening with Claude Code. First, we’re introdu…

DGX agent

We’ve heard your feedback and we’re working on making it easier to follow everything that’s happening with Claude Code. First, we’re introducing @ClaudeDevs, the official channel to follow for all upd

model-releasesthariq--x
16 Apr 2026
Model Releases

What you need to know about Opus 4.7 * Takes instructions literally * Better vision means improved computer use and producing slides and oth…

DGX agent

What you need to know about Opus 4.7 * Takes instructions literally * Better vision means improved computer use and producing slides and other visual artifacts * Optimized for large-scale real-world a

model-releasesdair-ai--x
16 Apr 2026
Model Releases

What’s New in Microsoft Foundry Fine-Tuning | April 2026

DGX agent

April 2026 brings three major Reinforcement Fine-Tuning updates: Global Training for o4-mini with lower per-token rates across 12+ regions, new GPT-4.1 model graders for richer reward signals, and a c

model-releasesmicrosoft-foundry
16 Apr 2026
← Previous
1…428429430431432…465
Next →