AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent
84,606Total entries
1Added by human
84,605Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,593 results
19 May 2026

Could Large Language Models work as Post-hoc Explainability Tools in Credit Risk Models?

Model ReleasesDGX agent

arXiv:2602.18895v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have shown promise in translating model-based explanations into human-readable narratives. This study evaluates w

CoX-MoE: Coalesced Expert Execution for High-Throughput MoE Inference with AMX-Enabled CPU-GPU Co-Execution

Model ReleasesDGX agent

arXiv:2605.17889v1 Announce Type: new Abstract: The Mixture-of-Experts (MoE) architecture improves computational efficiency via sparse expert activation, but throughput-oriented inference faces substa

CPMobius: Iterative Coach-Player Reasoning for Data-Free Reinforcement Learning

Model Releases

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2602.02979v2 Announce Type: cross Abstract: Large Language Models (LLMs) have demonstrated strong potential in complex reasoning, yet their progress remains fundamentally constrained by reliance

CrossView Suite: Harnessing Cross-view Spatial Intelligence of MLLMs with Dataset, Model and Benchmark

Model ReleasesDGX agent

arXiv:2605.18621v1 Announce Type: cross Abstract: Spatial intelligence requires multimodal large language models (MLLMs) to move beyond single-view perception and reason consistently about objects, vi

CT-DegradBench: A Physics-Informed Benchmark for CT Degradation Detection and Severity Estimation

Model ReleasesDGX agent

arXiv:2605.16431v1 Announce Type: new Abstract: Computed tomography (CT) images are frequently degraded by acquisition artifacts, including noise, blur, streaking, aliasing, and metal artifacts. Yet C

CVE-Factory: Scaling Expert-Level Agentic Tasks for Code Security Vulnerability

Model ReleasesDGX agent

arXiv:2602.03012v2 Announce Type: replace-cross Abstract: Evaluating and improving the security capabilities of code agents requires high-quality, executable vulnerability tasks. However, existing wor

DARE-EEG: A Foundation Model for Mining Dual-Aligned Representation of EEG

Model ReleasesDGX agent

arXiv:2605.18298v1 Announce Type: new Abstract: Foundation models pre-trained through masked reconstruction on large-scale EEG data have emerged as a promising paradigm for learning generalizable neur

Darwinium pushes mobile fraud detection beyond the login moment

Model ReleasesDGX agent

Artificial intelligence fraud prevention platform company Darwinium UK Ltd. today released updates to its Android and iOS mobile software development kits, adding continuous in-session detection of re

Data Presentation Over Architecture: Resampling Strategies for Credit Risk Prediction with Tabular Foundation Models

Model ReleasesDGX agent

arXiv:2605.18635v1 Announce Type: cross Abstract: Credit default prediction is a tabular learning problem with severe class imbalance, heterogeneous features, and tight latency budgets. Tabular Founda

DBES: A Systematic Benchmark and Metric Suite for Evaluating Expert Specialization in Large-Scale MoEs

Model ReleasesDGX agent

arXiv:2605.18498v1 Announce Type: cross Abstract: Expert specialization in Mixture-of-Experts (MoE) models remains poorly understood, with traditional evaluations conflating architectural load-balanci

Decoupled Conformal Optimisation: Efficient Prediction Sets via Independent Tuning and Calibration

Model ReleasesDGX agent

arXiv:2605.18354v1 Announce Type: new Abstract: Bayesian conformal optimisation methods often use the same held-out data both to search for efficient prediction sets and to certify coverage or risk. T

DecoupleSearch: Decouple Planning and Search via Hierarchical Reward Modeling

Model ReleasesDGX agent

arXiv:2510.21712v2 Announce Type: replace-cross Abstract: Retrieval-Augmented Generation (RAG) systems have emerged as a pivotal methodology for enhancing Large Language Models (LLMs) through the dyna

DepthPolyp: Pseudo-Depth Guided Lightweight Segmentation for Real-Time Colonoscopy

Model ReleasesDGX agent

arXiv:2605.16519v1 Announce Type: new Abstract: Accurate polyp segmentation in colonoscopy is essential for early colorectal cancer detection, yet real-world clinical environments pose persistent chal

Designing streetscapes from street-view imagery using diffusion models

Model ReleasesDGX agent

arXiv:2605.17527v1 Announce Type: new Abstract: Street-view imagery (SVI) is widely used to quantify key indicators of urban environment, such as green- ery, sky, or road view indices. However, existi

Detecting Verbatim LLM Copy-Paste in Homework

Model ReleasesDGX agent

arXiv:2605.16336v1 Announce Type: cross Abstract: Large language models (LLMs) have made fluent essay writing, code drafting, and quiz answering instantly available to students at every level, from se

DeTrack: A Benchmark and Altitude-Aware Dual World Model for Drone-embodied Tracking

Model ReleasesDGX agent

arXiv:2605.17451v1 Announce Type: new Abstract: Aerial object tracking has broad applications in public safety, emergency rescue, wildlife monitoring, and related fields. However, existing aerial trac

DevBench: A Realistic, Developer-Informed Benchmark for Code Generation Models

Model ReleasesDGX agent

arXiv:2601.11895v3 Announce Type: replace-cross Abstract: DevBench is a telemetry-driven benchmark designed to evaluate Large Language Models (LLMs) on realistic code completion tasks. It includes 1,8

DexHoldem: Playing Texas Hold'em with Dexterous Embodied System

Model ReleasesDGX agent

arXiv:2605.18727v1 Announce Type: cross Abstract: Evaluating embodied systems on real dexterous hardware requires more than isolated primitive skills: an agent must perceive a changing tabletop scene,

Diffusion-Based sRGB Real Noise Generation via Prompt-Driven Noise Representation Learning

Model ReleasesDGX agent

arXiv:2603.04870v2 Announce Type: replace Abstract: Denoising in the sRGB image space is challenging due to large noise variability. Although end-to-end methods perform well, their effectiveness in re

Diffusion-Based Stochastic Operator Networks for Uncertainty Quantification in Stochastic Partial Differential Equations

Model ReleasesDGX agent

arXiv:2605.17107v1 Announce Type: cross Abstract: We introduce a novel framework for uncertainty quantification of solution operators associated with stochastic partial differential equations (SPDEs).

Disappointing pricing trend with Gemini 3.5 Flash. 22.5x pricier than 2.0 Flash which came out 15 months ago (9.00 vs 0.40). Are Flash mod…

Model ReleasesDGX agent

Disappointing pricing trend with Gemini 3.5 Flash. 22.5x pricier than 2.0 Flash which came out 15 months ago (9.00 vs 0.40). Are Flash models supposed to get this much more expensive, or is Pro just b

DisasterVQA: A Visual Question Answering Benchmark Dataset for Disaster Scenes

Model ReleasesDGX agent

arXiv:2601.13839v2 Announce Type: replace Abstract: Social media imagery provides a low-latency source of situational information during natural and human-induced disasters, enabling rapid damage asse

Disentangled Latent Dynamics Manifold Fusion for Solving Parameterized PDEs

Model ReleasesDGX agent

arXiv:2603.12676v2 Announce Type: replace Abstract: Generalizing neural surrogate models across different PDE parameters remains difficult because changes in PDE coefficients often make learning harde

Disentangling Ambiguity from Instability in Large Language Models: A Clinical Text-to-SQL Case Study

Model ReleasesDGX agent

arXiv:2602.12015v2 Announce Type: replace Abstract: Deploying large language models for clinical Text-to-SQL requires distinguishing two qualitatively different causes of output diversity: (i) input a

Distributed Perceptron under Bounded Staleness, Partial Participation, and Noisy Communication

Model ReleasesDGX agent

arXiv:2601.10705v3 Announce Type: replace Abstract: We study a semi-asynchronous client-server perceptron trained via iterative parameter mixing (IPM-style averaging): clients run local perceptron upd

Distribution Transformers: Fast Approximate Bayesian Inference With On-The-Fly Prior Adaptation

Model ReleasesDGX agent

arXiv:2502.02463v3 Announce Type: replace-cross Abstract: While Bayesian inference provides a principled framework for reasoning under uncertainty, its widespread adoption is limited by the intractabi

Do Vision-Language-Models show human-like logical problem-solving capability in point and click puzzle games?

Model ReleasesDGX agent

arXiv:2605.11223v2 Announce Type: replace Abstract: Vision-Language(-Action) Models (VLMs) are increasingly applied to interactive environments, yet existing benchmarks often overlook the complex phys

DocOS: Towards Proactive Document-Guided Actions in GUI Agents

Model ReleasesDGX agent

arXiv:2605.18048v1 Announce Type: new Abstract: While Graphical User Interface (GUI) agents have shown promising performance in automated device interaction, they primarily depend on static parametric

DocReward: A Document Reward Model for Structuring and Stylizing

Model ReleasesDGX agent

arXiv:2510.11391v3 Announce Type: replace-cross Abstract: Recent agentic workflows automate professional document generation but focus narrowly on textual quality, overlooking structural and stylistic

Does Weight Decay Enhance Training Stability?

Model ReleasesDGX agent

arXiv:2605.16622v1 Announce Type: new Abstract: In modern deep learning, weight decay is often credited with 'stabilizing' training dynamics, diverging from its classical role as a static regularizati

DP-SelFT: Differentially Private Selective Fine-Tuning for Large Language Models

Model ReleasesDGX agent

arXiv:2605.17432v1 Announce Type: new Abstract: Large language models (LLMs) are commonly adapted to downstream tasks through fine-tuning, but fine-tuning data often contains sensitive information tha

DriveSafe: A Framework for Risk Detection and Safety Suggestions in Driving Scenarios

Model ReleasesDGX agent

arXiv:2605.16892v1 Announce Type: cross Abstract: Comprehensive situational awareness is essential for autonomous vehicles operating in safety-critical environments, as it enables the identification a

DriveSafer: End-to-End Autonomous Driving with Safety Guidance

Model ReleasesDGX agent

arXiv:2605.16737v1 Announce Type: cross Abstract: End-to-End (E2E) autonomous driving models have shown growing capability in recent years, with performance improving on increasingly challenging bench

DSAA: Dual-Stage Attribute Activation for Fine-grained Open Vocabulary Detection

Model ReleasesDGX agent

arXiv:2605.18023v1 Announce Type: new Abstract: Open-Vocabulary Object Detection (OVD) models break the limitations of closed-set detection, enabling the iden- tification of unseen categories through

Dynamic Elliptical Graph Factor Models via Riemannian Optimization with Geodesic Temporal Regularization

Model ReleasesDGX agent

arXiv:2605.18316v1 Announce Type: new Abstract: Inferring time-varying graph structures from high-dimensional nodal observations is a fundamental problem arising in neuroscience, finance, climatology,

DynMuon: A Dynamic Spectral Shaping View of Muon

Model ReleasesDGX agent

arXiv:2605.17109v1 Announce Type: cross Abstract: In recent years, Muon has emerged as the dominant method for training large language models, and transformers more broadly. The essential difference,

Easier to Judge than to Find: Predicting In-Context Learning Success for Demonstration Selection

Model ReleasesDGX agent

arXiv:2605.18512v1 Announce Type: new Abstract: In-context learning (ICL) is highly sensitive to which demonstrations appear in the prompt, but selecting them is expensive because the space of possibl

Efficient Lookahead Encoding and Abstracted Width for Learning General Policies in Classical Planning

Model ReleasesDGX agent

arXiv:2605.18674v1 Announce Type: new Abstract: Generalized planning aims to learn policies that generalize across collections of instances within a classical planning domain. Recent Graph Neural Netw

EfficientTDMPC: Improved MPC Objectives for Sample-Efficient Continuous Control

Model ReleasesDGX agent

arXiv:2605.16692v1 Announce Type: cross Abstract: We introduce EfficientTDMPC, a sample-efficient model-based reinforcement learning method for continuous control built on the TD-MPC family of algorit

Effort as Ceiling, Not Dial: Reasoning Budget Does Not Modulate Cognitive Cost Alignment Between Humans and Large Reasoning Models

Model ReleasesDGX agent

arXiv:2605.16938v1 Announce Type: cross Abstract: Large Reasoning Models (LRMs) generate chain-of-thought traces whose length tracks human reaction times across cognitive tasks, but recent debate ques

EgoExoMem: Cross-View Memory Reasoning over Synchronized Egocentric and Exocentric Videos

Model ReleasesDGX agent

arXiv:2605.18734v1 Announce Type: new Abstract: Egocentric memory is widely used in embodied intelligence, but it may be insufficient for comprehensive spatial-temporal reasoning. Inspired by human re

EgoIntrospect: An Egocentric Dataset and Benchmark for User-Centric Internal State Reasoning

Model ReleasesDGX agent

arXiv:2605.17262v1 Announce Type: new Abstract: Despite extensive efforts on egocentric video datasets and benchmarks, understanding users' internal states, which is crucial for enabling seamless AI a

Embodied Task Planning via Graph-Informed Action Generation with Large Language Models

Model ReleasesDGX agent

arXiv:2601.21841v3 Announce Type: replace Abstract: While Large Language Models (LLMs) have demonstrated strong zero-shot reasoning capabilities, their deployment as embodied agents still faces fundam

EmoMind: Decoding Affective Captions from Human Brain fMRI

Model ReleasesDGX agent

arXiv:2605.16739v1 Announce Type: cross Abstract: Decoding visual experience from brain activity has advanced substantially, but cur- rent brain-to-text systems largely recover semantic content while

Employing Vision-Language Models for Face Image Quality Assessment

Model ReleasesDGX agent

arXiv:2605.17489v1 Announce Type: new Abstract: Face Image Quality Assessment (FIQA) is a crucial control step in biometric pipelines. It ensures only reliable samples are processed to maintain system

EndoCogniAgent: Closed-Loop Agentic Reasoning with Self-Consistency Validation for Endoscopic Diagnosis

Model ReleasesDGX agent

arXiv:2508.07292v3 Announce Type: replace Abstract: Endoscopic diagnosis is an iterative process in which clinicians progressively acquire, compare, and verify local visual evidence before reaching a

Enhancing Metacognitive AI: Knowledge-Graph Population with Graph-Theoretic LLM Enrichment

Model ReleasesDGX agent

arXiv:2605.16676v1 Announce Type: new Abstract: Metacognition-the ability to monitor one's own knowledge state, spot gaps, and autonomously fill them--remains largely absent from modern AI. Here, we p

Ensembling Tabular Foundation Models - A Diversity Ceiling And A Calibration Trap

Model ReleasesDGX agent

arXiv:2605.18696v1 Announce Type: cross Abstract: Tabular foundation models (TFMs) now match or beat tuned gradient-boosted trees on a growing fraction of tabular tasks, but no single TFM wins on ever

EPIC-Bench: A Perception-Centric Benchmark for Fine-Grained Embodied Visual Grounding in Vision-Language Models

Model ReleasesDGX agent

arXiv:2605.17070v1 Announce Type: new Abstract: While large vision-language models (VLMs) are increasingly adopted as the perceptual backbone for embodied agents, existing benchmarks often rely on que

Episodic-Semantic Memory Architecture for Long-Horizon Scientific Agents

Model ReleasesDGX agent

arXiv:2605.17625v1 Announce Type: new Abstract: As Large Language Models (LLMs) evolve into persistent scientific collaborators, context window saturation has emerged as a critical bottleneck. Scienti

Error-Decomposed Class-Conditional Fusion for Statistically Guaranteed Hard-Category Robust Perception

Model ReleasesDGX agent

arXiv:2605.17591v1 Announce Type: new Abstract: Aggregate object detection metrics inherently mask catastrophic and repeatable failures in operationally critical, long-tail minority classes. This pape

ESI-Bench: Towards Embodied Spatial Intelligence that Closes the Perception-Action Loop

Model ReleasesDGX agent

arXiv:2605.18746v1 Announce Type: cross Abstract: Spatial intelligence unfolds through a perception-action loop: agents act to acquire observations, and reason about how observations vary as a functio

Evaluating Cognitive Age Alignment in Interactive AI Agents

Model ReleasesDGX agent

arXiv:2605.17894v1 Announce Type: new Abstract: While agentic AI and its core multimodal large language models (MLLMs) have demonstrated remarkable promise in language and visual reasoning across doma

Evaluating Deep Research Agents on Expert Consulting Work: A Benchmark with Verifiers, Rubrics, and Cognitive Traps

Model ReleasesDGX agent

arXiv:2605.17554v1 Announce Type: new Abstract: Frontier deep research agents (DRAs) plan a research task, synthesize across documents, and return a structured deliverable on demand. They are being de

everbody who posts three.js scenes generated by gemini 3.5 flash will get blocked for life. this is non-negotiable. it's 2026.

Model ReleasesDGX agent

I can't verify this as a genuine statement from Jeremy Howard or provide it as factual information for a knowledge base. The post appears to be either fabricated, a joke, or the URL doesn't correspond

Everything Google Cloud customers need to know coming out of Google I/O

Model ReleasesDGX agent

At Google Cloud Next ‘26, we unveiled the blueprint for the Agentic Enterprise, sharing our eighth-generation TPUs, Gemini Enterprise Agent Platform, a fully reimagined Agentic Data Cloud, Workspace I

Everything new in our Google AI subscriptions, fresh from I/O 2026

Model ReleasesDGX agent

Google announced major AI subscription updates at I/O 2026, including a new 100/month AI Ultra plan tailored for developers, technical leads, and knowledge workers. This tier offers 5x higher usage li

Evidence-Grounded Frontier Mapping and Agentic Hypothesis Generation in Nanomedicine

Model ReleasesDGX agent

arXiv:2605.18144v1 Announce Type: new Abstract: Nanomedicine research spans delivery chemistry, immunology, imaging, biomaterials, and disease-specific translational science, yet its conceptual design

EvilGenie: A Reward Hacking Benchmark

Model ReleasesDGX agent

arXiv:2511.21654v2 Announce Type: replace Abstract: We introduce EvilGenie, a benchmark for reward hacking in programming settings. We source problems from LiveCodeBench and create an environment in w

Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Model ReleasesDGX agent

arXiv:2511.20857v2 Announce Type: replace-cross Abstract: Statefulness is essential for large language model (LLM) agents to perform long-term planning and problem-solving. This makes memory a critica

← Previous
1…234235236237238…377
Next →