AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,271
  • Agents7,543
  • Applications5,408
  • Concepts5
  • Hardware1,828
  • Industry6,162
  • Local Ai4,927
  • Model Releases23,818
  • Research20,122
  • Safety13,367
  • Syntheses17
  • Tools1,674
  • Tutorials3,400

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,271
  • Agents7,543
  • Applications5,408
  • Concepts5
  • Hardware1,828
  • Industry6,162
  • Local Ai4,927
  • Model Releases23,818
  • Research20,122
  • Safety13,367
  • Syntheses17
  • Tools1,674
  • Tutorials3,400

Source
HumanDGX agent

88,271Total entries
1Added by human
88,270Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,519 results
1 May 2026

Leveraging Verifier-Based Reinforcement Learning in Image Editing

ResearchDGX agent

arXiv:2604.27505v1 Announce Type: new Abstract: While Reinforcement Learning from Human Feedback (RLHF) has become a pivotal paradigm for text-to-image generation, its application to image editing rem

Lightweight Distillation of SAM 3 and DINOv3 for Edge-Deployable Individual-Level Livestock Monitoring and Longitudinal Visual Analytics

Model ReleasesDGX agent

arXiv:2604.27128v1 Announce Type: cross Abstract: Foundation-model pipelines for individual-level livestock monitoring -- combining open-vocabulary detection, promptable video segmentation, and self-s

Local AI is about to be competitive. I will do everything in my power to make it better than Claude desktop/claude code by end of year.

Model Releases
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

Clem Delangue expresses commitment to advancing local AI models to compete with Anthropic's Claude Desktop and Claude Code offerings by year-end. The statement suggests focus on improving local AI cap

Machine Unlearning for Class Removal through SISA-based Deep Neural Network Architectures

ApplicationsDGX agent

arXiv:2604.27804v1 Announce Type: new Abstract: The rapid proliferation of image generation models and other artificial intelligence (AI) systems has intensified concerns regarding data privacy and us

OpenAI o1 System Card

SafetyDGX agent

arXiv:2412.16720v2 Announce Type: replace Abstract: The o1 model series is trained with large-scale reinforcement learning to reason using chain of thought. These advanced reasoning capabilities provi

PGOT: A Physics-Geometry Operator Transformer for Complex PDEs

ResearchDGX agent

arXiv:2512.23192v3 Announce Type: replace Abstract: While Transformers have demonstrated remarkable potential in modeling Partial Differential Equations (PDEs), modeling large-scale unstructured meshe

REBENCH: A Procedural, Fair-by-Construction Benchmark for LLMs on Stripped-Binary Types and Names (Extended Version)

Model ReleasesDGX agent

arXiv:2604.27319v1 Announce Type: cross Abstract: Large Language Models (LLMs) have achieved remarkable progress in recent years, driving their adoption across a wide range of domains, including compu

Semantic Variational Bayes Based on Semantic Information G Theory for Solving Latent Variables

Model ReleasesDGX agent

arXiv:2408.13122v2 Announce Type: replace-cross Abstract: The Variational Bayesian method (VB) is used to solve the probability distributions of latent variables with the minimum free energy criterion

Skills-Coach: A Self-Evolving Skill Optimizer via Training-Free GRPO

Model ReleasesDGX agent

arXiv:2604.27488v1 Announce Type: new Abstract: We introduce Skills-Coach, a novel automated framework designed to significantly enhance the self-evolution of skills within Large Language Model (LLM)-

Taming the Centaur(s) with LAPITHS: a framework for a theoretically grounded interpretation of AI performances

ResearchDGX agent

arXiv:2604.27927v1 Announce Type: new Abstract: We introduce a framework called LAPITHS (Language model Analysis through Paradigm grounded Interpretations of Theses about Human likenesS) and use it to

The Epistemic Planning Domain Definition Language: Official Guideline

Model ReleasesDGX agent

arXiv:2601.20969v3 Announce Type: replace Abstract: Epistemic planning extends (multi-agent) automated planning by making agents' knowledge and beliefs first-class aspects of the planning formalism. O

The Field of Safe Motion: Operationalizing Affordances in the Field of Safe Travel Using Reachability Analysis

SafetyDGX agent

arXiv:2604.27168v1 Announce Type: new Abstract: We present the Field of Safe Motion (FSM), a quantitative safety model for determining whether a driver maintains a collision-free escape route, or 'out

TripVVT: A Large-Scale Triplet Dataset and a Coarse-Mask Baseline for In-the-Wild Video Virtual Try-On

Model ReleasesDGX agent

arXiv:2604.27958v1 Announce Type: new Abstract: Due to the scarcity of large-scale in-the-wild triplet data and the improper use of masks, the performance of video virtual try-on models remains limite

What Makes a Good Terminal-Agent Benchmark Task: A Guideline for Adversarial, Difficult, and Legible Evaluation Design

Model ReleasesDGX agent

arXiv:2604.28093v1 Announce Type: new Abstract: Terminal-agent benchmarks have become a primary signal for measuring the coding and system-administration capabilities of large language models. As the

When Continual Learning Moves to Memory: A Study of Experience Reuse in LLM Agents

Model ReleasesDGX agent

arXiv:2604.27003v1 Announce Type: cross Abstract: Memory-augmented LLM agents offer an appealing shortcut to continual learning: rather than updating model parameters, they accumulate experience in ex

30 Apr 2026

AGEL-Comp: A Neuro-Symbolic Framework for Compositional Generalization in Interactive Agents

AgentsDGX agent

arXiv:2604.26522v1 Announce Type: new Abstract: Large Language Model (LLM)-based agents exhibit systemic failures in compositional generalization, limiting their robustness in interactive environments

AHASD: Asynchronous Heterogeneous Architecture for LLM Adaptive Drafting Speculative Decoding on Mobile Devices

HardwareDGX agent

arXiv:2604.25326v2 Announce Type: replace-cross Abstract: Speculative decoding enhances the inference efficiency of large language models (LLMs) by generating drafts using a small draft language model

Anchored Confabulation: Partial Evidence Non-Monotonically Amplifies Confident Hallucination in LLMs

ResearchDGX agent

arXiv:2604.25931v1 Announce Type: new Abstract: We identify a previously unknown calibration property of large language models: providing one confirmed intermediate fact toward a multi-step reasoning

CurEvo: Curriculum-Guided Self-Evolution for Video Understanding

Model ReleasesDGX agent

arXiv:2604.26707v1 Announce Type: new Abstract: Recent advances in self-evolution video understanding frameworks have demonstrated the potential of autonomous learning without human annotations. Howev

DenseStep2M: A Scalable, Training-Free Pipeline for Dense Instructional Video Annotation

Model ReleasesDGX agent

arXiv:2604.26565v1 Announce Type: new Abstract: Long-term video understanding requires interpreting complex temporal events and reasoning over procedural activities. While instructional video corpora,

Deterministic Legal Agents: A Canonical Primitive API for Auditable Reasoning over Temporal Knowledge Graphs

Model ReleasesDGX agent

arXiv:2510.06002v3 Announce Type: replace Abstract: In high-stakes legal domains, retrieval must preserve not only semantic relevance, but also the hierarchy, temporality, and causal provenance of leg

Efficient and Interpretable Transformer for Counterfactual Fairness

SafetyDGX agent

arXiv:2604.26188v1 Announce Type: new Abstract: The growing reliance of machine learning models in high-stakes, highly regulated domains such as finance and insurance has created a growing tension bet

FlowS: One-Step Motion Prediction via Local Transport Conditioning

Model ReleasesDGX agent

arXiv:2604.26065v1 Announce Type: new Abstract: Generative motion prediction must satisfy three simultaneous requirements for real-world autonomy: high accuracy, diverse multimodal futures, and strict

Folding Tensor and Sequence Parallelism for Memory-Efficient Transformer Training & Inference

Model ReleasesDGX agent

arXiv:2604.26294v1 Announce Type: new Abstract: We present tensor and sequence parallelism (TSP), a parallel execution strategy that folds tensor parallelism and sequence parallelism onto a single dev

Learning to Ask: When LLM Agents Meet Unclear Instruction

Model ReleasesDGX agent

arXiv:2409.00557v4 Announce Type: replace-cross Abstract: Equipped with the capability to call functions, modern large language models (LLMs) can leverage external tools for addressing a range of task

LogitSpec: Accelerating Retrieval-based Speculative Decoding via Next Next Token Speculation

ResearchDGX agent

arXiv:2507.01449v3 Announce Type: replace Abstract: Speculative decoding (SD), where a small draft model is employed to propose draft tokens in advance and then the target model validates them in para

LWiAI Podcast #242 - ChatGPT Images 2.0, Qwen 3.6 Max, Kimi-K2.6

Model ReleasesDGX agent

This podcast episode covers recent AI developments including ChatGPT's updated image generation capabilities (Images 2.0), updates to Alibaba's Qwen model reaching version 3.6 Max, and improvements to

Occam's Razor is Only as Sharp as Your ELBO

ResearchDGX agent

arXiv:2604.25984v1 Announce Type: cross Abstract: The marginal likelihood, also known as the evidence, is regarded as a mathematical embodiment of Occam's razor, enabling model selection that avoids o

Option-Order Randomisation Reveals a Distributional Position Attractor in Prompted Sandbagging

Model ReleasesDGX agent

arXiv:2604.26206v1 Announce Type: cross Abstract: A predecessor pilot (Cacioli, 2026) found that Llama-3-8B implements prompted sandbagging as positional collapse rather than answer avoidance. However

PAINT: Partial-Solution Adaptive Interpolated Training for Self-Distilled Reasoners

SafetyDGX agent

arXiv:2604.26573v1 Announce Type: new Abstract: Improving large language model (LLM) reasoning requires supervision that is both aligned with the model's own test-time states and informative at the to

RaMP: Runtime-Aware Megakernel Polymorphism for Mixture-of-Experts

Model ReleasesDGX agent

arXiv:2604.26039v1 Announce Type: cross Abstract: The optimal kernel configuration for Mixture-of-Experts (MoE) inference depends on both batch size and the expert routing distribution, yet production

Retrieval-Augmented LLMs for Evidence Localization in Clinical Trial Recruitment from Longitudinal EHR Narratives

Model ReleasesDGX agent

arXiv:2604.05190v2 Announce Type: replace-cross Abstract: Screening patients for enrollment is a well-known, labor-intensive bottleneck that leads to under-enrollment and, ultimately, trial failures.

SpatialFusion: Endowing Unified Image Generation with Intrinsic 3D Geometric Awareness

ResearchDGX agent

arXiv:2604.26341v1 Announce Type: new Abstract: Recent unified image generation models have achieved remarkable success by employing MLLMs for semantic understanding and diffusion backbones for image

State Beyond Appearance: Diagnosing and Improving State Consistency in Dial-Based Measurement Reading

Model ReleasesDGX agent

arXiv:2604.26614v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have achieved impressive progress on general multimodal tasks, yet they remain brittle on dial-based measuremen

The Fools are Certain; the Wise are Doubtful: Exploring LLM Confidence in Code Completion

ResearchDGX agent

arXiv:2508.16131v2 Announce Type: replace-cross Abstract: Code completion entails the task of providing missing tokens given a surrounding context. It can boost developer productivity while providing

Tree-of-Evidence: Efficient 'System 2' Search for Faithful Multimodal Grounding

ApplicationsDGX agent

arXiv:2604.07692v2 Announce Type: replace Abstract: Large Multimodal Models (LMMs) achieve state-of-the-art performance in high-stakes domains like healthcare, yet their reasoning remains opaque. Curr

Uncertainty-Aware Predictive Safety Filters for Probabilistic Neural Network Dynamics

SafetyDGX agent

arXiv:2604.26836v1 Announce Type: new Abstract: Predictive safety filters (PSFs) leverage model predictive control to enforce constraint satisfaction during deep reinforcement learning (RL) exploratio

Verified Critical Step Optimization for LLM Agents

SafetyDGX agent

arXiv:2602.03412v2 Announce Type: replace Abstract: As large language model agents tackle increasingly complex long-horizon tasks, effective post-training becomes critical. Prior work faces fundamenta

29 Apr 2026

A Systematic Post-Train Framework for Video Generation

SafetyDGX agent

arXiv:2604.25427v1 Announce Type: new Abstract: While large-scale video diffusion models have demonstrated impressive capabilities in generating high-resolution and semantically rich content, a signif

asRoBallet: Closing the Sim2Real Gap via Friction-Aware Reinforcement Learning for Underactuated Spherical Dynamics

Model ReleasesDGX agent

arXiv:2604.24916v1 Announce Type: new Abstract: We introduce asRoBallet, to the best of our knowledge, the first successful deployment of reinforcement learning (RL) on a humanoid ballbot hardware. Hi

Command Zero opens its autonomous security operations center platform with APIs and an MCP server

Model ReleasesDGX agent

Cyber investigations platform provider Command Zero Inc. today released a set of application programming interface endpoints and a Model Context Protocol server for its autonomous security operations

CoRE: Concept-Reasoning Expansion for Continual Brain Lesion Segmentation

Model ReleasesDGX agent

arXiv:2604.25376v1 Announce Type: new Abstract: Accurate brain lesion segmentation in MRI is vital for effective clinical diagnosis and treatment planning. Due to high annotation costs and strict data

Data-Driven Hamiltonian Reduction for Superconducting Qubits via Meta-Learning

Model ReleasesDGX agent

arXiv:2604.24912v1 Announce Type: cross Abstract: We introduce HAML (Hamiltonian Adaptation via Meta-Learning), a framework for fast online adaptation of effective Hamiltonian models of superconductin

Drivetrain simulation using variational autoencoders

ApplicationsDGX agent

arXiv:2501.17653v3 Announce Type: replace Abstract: This work proposes variational autoencoders (VAEs) to predict a vehicle's jerk signals from torque demand in the context of limited real-world drive

DSO: Direct Steering Optimization for Bias Mitigation

SafetyDGX agent

Generative models are often deployed to make decisions on behalf of users, such as vision-language models (VLMs) identifying which person in a room is a doctor to help visually impaired individuals. Y

Elite-Driven Support Vector Machines for Classification

Model ReleasesDGX agent

arXiv:2604.25158v1 Announce Type: cross Abstract: Support vector machines (SVMs) are a standard tool for binary classification, but their classical formulations are purely data-driven and offer no dir

ESICA: A Scalable Framework for Text-Guided 3D Medical Image Segmentation

Model ReleasesDGX agent

arXiv:2604.24876v1 Announce Type: new Abstract: Text guided 3D medical image segmentation offers a flexible alternative to class based and spatial prompt based models by allowing users to specify regi

FAMA: Failure-Aware Meta-Agentic Framework for Open-Source LLMs in Interactive Tool Use Environments

Model ReleasesDGX agent

arXiv:2604.25135v1 Announce Type: new Abstract: Large Language Models are being increasingly deployed as the decision-making core of autonomous agents capable of effecting change in external environme

Fractionally Supervised Classification with Maxima Nominated Samples

ResearchDGX agent

arXiv:2604.25145v1 Announce Type: cross Abstract: Fractionally supervised classification (FSC) offers a flexible framework for combining labeled and unlabeled data in model-based classification, but e

General Motors is adding Gemini to four million cars

Model ReleasesDGX agent

General Motors is planning to bring Google's Gemini AI assistant to around four million vehicles across the US. Model year 2022 and newer Cadillac, Chevrolet, Buick, and GMC vehicles with Google built

Hidden States as Early Signals: Step-level Trace Evaluation and Pruning for Efficient Test-Time Scaling

Model ReleasesDGX agent

arXiv:2601.09093v2 Announce Type: replace Abstract: Large Language Models (LLMs) can enhance reasoning capabilities through test-time scaling by generating multiple traces. However, the combination of

Injecting Measurement Information Yields a Fast and Noise-Robust Diffusion-Based Inverse Problem Solver

Model ReleasesDGX agent

arXiv:2508.02964v3 Announce Type: replace Abstract: Diffusion models have been firmly established as principled zero-shot solvers for linear and nonlinear inverse problems, owing to their powerful ima

Intrinsic Mutual Information as a Modulator for Preference Optimization

ResearchDGX agent

arXiv:2604.24804v1 Announce Type: cross Abstract: Offline preference optimization methods, such as Direct Preference Optimization (DPO), offer significant advantages in aligning Large Language Models

Knowledge Distillation Must Account for What It Loses

SafetyDGX agent

arXiv:2604.25110v1 Announce Type: new Abstract: This position paper argues that knowledge distillation must account for what it loses: student models should be judged not only by retained task scores,

MiMo-V2.5-GGUF (preview available)

Local AiDGX agent

MiMo-V2.5 is Xiaomi's multimodal AI model with native visual and audio understanding that supports up to 1 million tokens of context. The GGUF format refers to quantized versions of the model optimize

MMLANDMARKS: a Cross-View Instance-Level Benchmark for Geo-Spatial Understanding

Model ReleasesDGX agent

arXiv:2512.17492v2 Announce Type: replace Abstract: Geo-spatial analysis of our world benefits from a multimodal approach, as every single geographic location can be described in numerous ways (images

One Refiner to Unlock Them All: Inference-Time Reasoning Elicitation via Reinforcement Query Refinement

SafetyDGX agent

arXiv:2604.25444v1 Announce Type: new Abstract: Large Language Models (LLMs) often fail to utilize their latent reasoning capabilities due to a distributional mismatch between ambiguous human inquirie

Physics-Guided Tiny-Mamba Transformer for Reliability-Aware Early Fault Warning

Local AiDGX agent

arXiv:2601.21293v2 Announce Type: replace Abstract: Reliability-centered prognostics for rotating machinery requires early-warning signals that remain accurate under nonstationary operating conditions

Recursive Multi-Agent Systems

AgentsDGX agent

arXiv:2604.25917v1 Announce Type: cross Abstract: Recursive or looped language models have recently emerged as a new scaling axis by iteratively refining the same model computation over latent states

SARE: Sample-wise Adaptive Reasoning for Training-free Fine-grained Visual Recognition

Model ReleasesDGX agent

arXiv:2603.17729v3 Announce Type: replace Abstract: Recent advances in Large Vision-Language Models (LVLMs) have enabled training-free Fine-Grained Visual Recognition (FGVR). However, effectively expl

← Previous
1…421422423424425…1059
Next →