AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
48,975 results
Model Releases

OphMAE: Bridging Volumetric and Planar Imaging with a Foundation Model for Adaptive Ophthalmological Diagnosis

DGX agent

arXiv:2605.02714v1 Announce Type: new Abstract: The advent of foundation models has heralded a new era in medical artificial intelligence (AI), enabling the extraction of generalizable representations

model-releasesarxiv-cs-cv
5 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Orthographic Constraint Satisfaction and Human Difficulty Alignment in Large Language Models

DGX agent

arXiv:2511.21086v2 Announce Type: replace Abstract: Large language models must satisfy hard orthographic constraints during controlled text generation, yet systematic cross-family evaluation remains l

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Pretraining A Large Language Model using Distributed GPUs: A Memory-Efficient Decentralized Paradigm

DGX agent

arXiv:2602.11543v2 Announce Type: replace Abstract: Pretraining large language models (LLMs) typically requires centralized clusters with thousands of high-memory GPUs (e.g., H100/A100). Recent decent

model-releasesarxiv-cs-cl
5 May 2026
Research

Consistent Diffusion Language Models

DGX agent

arXiv:2605.00161v1 Announce Type: new Abstract: Diffusion language models (DLMs) are an attractive alternative to autoregressive models because they promise sublinear-time, parallel generation, yet pr

researcharxiv-cs-lg
4 May 2026
Local Ai

RadLite: Multi-Task LoRA Fine-Tuning of Small Language Models for CPU-Deployable Radiology AI

DGX agent

arXiv:2605.00421v1 Announce Type: new Abstract: Large language models (LLMs) show promise in radiology but their deployment is limited by computational requirements that preclude use in resource-const

local-aiarxiv-cs-cl
4 May 2026
Research

The Silent Thought: Modeling Internal Cognition in Full-Duplex Spoken Dialogue Models via Latent Reasoning

DGX agent

arXiv:2603.17837v3 Announce Type: replace-cross Abstract: During conversational interactions, humans subconsciously engage in concurrent thinking while listening to a speaker. Although this internal c

researcharxiv-cs-cl
4 May 2026
Model Releases

CL-bench Life: Can Language Models Learn from Real-Life Context?

DGX agent

arXiv:2604.27043v1 Announce Type: new Abstract: Today's AI assistants such as OpenClaw are designed to handle context effectively, making context learning an increasingly important capability for mode

model-releasesarxiv-cs-cl
1 May 2026
Research

Efficient-DLM: From Autoregressive to Diffusion Language Models, and Beyond in Speed

DGX agent

arXiv:2512.14067v2 Announce Type: replace-cross Abstract: Diffusion language models (dLMs) have emerged as a promising paradigm that enables parallel, non-autoregressive generation, but their learning

researcharxiv-cs-ai
1 May 2026
Model Releases

HERMES++: Toward a Unified Driving World Model for 3D Scene Understanding and Generation

DGX agent

arXiv:2604.28196v1 Announce Type: new Abstract: Driving world models serve as a pivotal technology for autonomous driving by simulating environmental dynamics. However, existing approaches predominant

model-releasesarxiv-cs-cv
1 May 2026
Agents

Heterogeneous Scientific Foundation Model Collaboration

DGX agent

arXiv:2604.27351v1 Announce Type: new Abstract: Agentic large language model systems have demonstrated strong capabilities. However, their reliance on language as the universal interface fundamentally

agentsarxiv-cs-ai
1 May 2026
Model Releases

Useless but Safe? Benchmarking Utility Recovery with User Intent Clarification in Multi-Turn Conversations

DGX agent

arXiv:2604.27093v1 Announce Type: cross Abstract: Current LLM safety alignment techniques improve model robustness against adversarial attacks, but overlook whether and how LLMs can recover helpfulnes

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

Visual Generation in the New Era: An Evolution from Atomic Mapping to Agentic World Modeling

DGX agent

arXiv:2604.28185v1 Announce Type: new Abstract: Recent visual generation models have made major progress in photorealism, typography, instruction following, and interactive editing, yet they still str

model-releasesarxiv-cs-cv
1 May 2026
Applications

When Your LLM Reaches End-of-Life: A Framework for Confident Model Migration in Production Systems

DGX agent

arXiv:2604.27082v1 Announce Type: new Abstract: We present a framework for migrating production Large Language Model (LLM) based systems when the underlying model reaches end-of-life or requires repla

applicationsarxiv-cs-ai
1 May 2026
Model Releases

Cross-Domain Transfer of Hyperspectral Foundation Models

DGX agent

arXiv:2604.26478v1 Announce Type: new Abstract: Hyperspectral imaging (HSI) semantic segmentation typically relies on in-domain training, but limited data availability often restricts model performanc

model-releasesarxiv-cs-cv
30 Apr 2026
Model Releases

LLM Psychosis: A Theoretical and Diagnostic Framework for Reality-Boundary Failures in Large Language Models

DGX agent

arXiv:2604.25934v1 Announce Type: cross Abstract: The deployment of large language models (LLMs) as interactive agents has exposed a category of behavioral failure that prevailing terminology, princip

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

Time Blindness: Why Video-Language Models Can't See What Humans Can?

DGX agent

arXiv:2505.24867v2 Announce Type: replace-cross Abstract: Recent advances in vision-language models (VLMs) have made impressive strides in understanding spatio-temporal relationships in videos. Howeve

model-releasesarxiv-cs-ai
30 Apr 2026
Safety

Beyond Accuracy: Benchmarking Cross-Task Consistency in Unified Multimodal Models

DGX agent

arXiv:2604.25072v1 Announce Type: new Abstract: Unified Multimodal Models (uMMs) aim to support both visual understanding and visual generation within a shared representation. However, existing evalua

safetyarxiv-cs-cv
29 Apr 2026
Model Releases

Evaluating Computational Pathology Foundation Models for Prostate Cancer Grading under Distribution Shifts

DGX agent

arXiv:2410.06723v2 Announce Type: replace-cross Abstract: Pathology foundation models (PFMs) have emerged as powerful pretrained encoders for computational pathology, but their robustness under clinic

model-releasesarxiv-cs-cv
29 Apr 2026
Tutorials

Exploring Time Conditioning in Diffusion Generative Models from Disjoint Noisy Data Manifolds

DGX agent

arXiv:2604.25289v1 Announce Type: cross Abstract: Practically, training diffusion models typically requires explicit time conditioning to guide the network through the denoising sampling process. Espe

tutorialsarxiv-cs-cv
29 Apr 2026
Research

Making AI-Assisted Grant Evaluation Auditable without Exposing the Model

DGX agent

arXiv:2604.25200v1 Announce Type: cross Abstract: Public agencies are beginning to consider large language models (LLMs) as decision-support tools for grant evaluation. This creates a practical govern

researcharxiv-cs-lg
29 Apr 2026
Research

Principled Detection of Hallucinations in Large Language Models via Multiple Testing

DGX agent

arXiv:2508.18473v3 Announce Type: replace Abstract: While Large Language Models (LLMs) have emerged as powerful foundational models to solve a variety of tasks, they have also been shown to be prone t

researcharxiv-cs-cl
29 Apr 2026
Applications

Robustness Evaluation of a Foundation Segmentation Model Under Simulated Domain Shifts in Abdominal CT: Implications for Health Digital Twin Deployment

DGX agent

arXiv:2604.25685v1 Announce Type: cross Abstract: Foundation segmentation models such as the Segment Anything Model (SAM) have demonstrated strong generalization across natural images; however, their

applicationsarxiv-cs-cv
29 Apr 2026
Research

Sketch2Arti: Sketch-based Articulation Modeling of CAD Objects

DGX agent

arXiv:2604.25781v1 Announce Type: new Abstract: Articulation modeling aims to infer movable parts and their motion parameters for a 3D object, enabling interactive animation, simulation, and shape edi

researcharxiv-cs-cv
29 Apr 2026
Hardware

Tendon-Actuated Robots with a Tapered, Flexible Polymer Backbone: Design, Fabrication, and Modeling

DGX agent

arXiv:2603.19124v2 Announce Type: replace Abstract: This paper presents the design, modeling, and fabrication of 3D-printed, tendon-actuated continuum robots featuring a flexible, tapered backbone con

hardwarearxiv-cs-ro
29 Apr 2026
Model Releases

Domain-Adapted Fine-Tuning of ECG Foundation Models for Multi-Label Structural Heart Disease Screening

DGX agent

arXiv:2604.23385v1 Announce Type: new Abstract: Transthoracic echocardiography is the reference standard for confirming structural heart disease (SHD), but first-line screening is limited by cost, wor

model-releasesarxiv-cs-lg
28 Apr 2026
Model Releases

EmoBench-M: Benchmarking Emotional Intelligence for Multimodal Large Language Models

DGX agent

arXiv:2502.04424v4 Announce Type: replace-cross Abstract: With the integration of multimodal large language models (MLLMs) into robotic systems and AI applications, embedding emotional intelligence (E

model-releasesarxiv-cs-ai
28 Apr 2026
Research

Impact of Age Specialized Models for Hypoglycemia Classification

DGX agent

arXiv:2604.23732v1 Announce Type: cross Abstract: Disease progression varies with age and is influenced by underlying genetic, biochemical, and hormonal etiologies, suggesting the need for tailored mo

researcharxiv-cs-ai
28 Apr 2026
Research

KERV: Kinematic-Rectified Speculative Decoding for Embodied VLA Models

DGX agent

arXiv:2603.01581v2 Announce Type: replace-cross Abstract: Vision-Language-Action (VLA) models build a token-domain robot control paradigm, yet suffer from low speed. Speculative Decoding (SD) is an op

researcharxiv-cs-lg
28 Apr 2026
Model Releases

Layerwise Convergence Fingerprints for Runtime Misbehavior Detection in Large Language Models

DGX agent

arXiv:2604.24542v1 Announce Type: cross Abstract: Large language models deployed at runtime can misbehave in ways that clean-data validation cannot anticipate: training-time backdoors lie dormant unti

model-releasesarxiv-cs-ai
28 Apr 2026
Applications

M^2-VLA: Boosting Vision-Language Models for Generalizable Manipulation via Layer Mixture and Meta-Skills

DGX agent

arXiv:2604.24182v1 Announce Type: new Abstract: Current Vision-Language-Action (VLA) models predominantly rely on end-to-end fine-tuning. While effective, this paradigm compromises the inherent genera

applicationsarxiv-cs-ro
28 Apr 2026
Model Releases

PivotMerge: Bridging Heterogeneous Multimodal Pre-training via Post-Alignment Model Merging

DGX agent

arXiv:2604.22823v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) rely on multimodal pre-training over diverse data sources, where different datasets often induce complementar

model-releasesarxiv-cs-ai
28 Apr 2026
Research

Using Language Models as Closed-Loop High-Level Planners for Robotics Applications: A Brief Overview and Benchmarks

DGX agent

arXiv:2511.07410v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) and Vision Language Models (VLMs) have become popular tools for embodied high-level planning. However, their depl

researcharxiv-cs-ai
28 Apr 2026
Model Releases

V-SEAM: Visual Semantic Editing and Attention Modulating for Causal Interpretability of Vision-Language Models

DGX agent

arXiv:2509.14837v2 Announce Type: replace Abstract: Recent advances in causal interpretability have extended from language models to vision-language models (VLMs), seeking to reveal their internal mec

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

VAPO: End-to-end Slide-Enhanced Speech Recognition with Omni-modal Large Language Models

DGX agent

arXiv:2510.08618v2 Announce Type: replace-cross Abstract: Omni-modal large language models (OLLMs) offer a promising end-to-end solution for slide-enhanced speech recognition due to their inherent mul

model-releasesarxiv-cs-cv
28 Apr 2026
Local Ai

Agentic World Modeling: Foundations, Capabilities, Laws, and Beyond

DGX agent

arXiv:2604.22748v1 Announce Type: new Abstract: As AI systems move from generating text to accomplishing goals through sustained interaction, the ability to model environment dynamics becomes a centra

local-aiarxiv-cs-ai
27 Apr 2026
Model Releases

Calibrating Behavioral Parameters with Large Language Models

DGX agent

arXiv:2602.01022v2 Announce Type: replace-cross Abstract: Behavioral parameters such as loss aversion, herding, and extrapolation are central to asset pricing models but remain difficult to measure re

model-releasesarxiv-cs-ai
27 Apr 2026
Research

Large Language Models Decide Early and Explain Later

DGX agent

arXiv:2604.22266v1 Announce Type: new Abstract: Large Language Models often achieve strong performance by generating long intermediate chain-of-thought reasoning. However, it remains unclear when a mo

researcharxiv-cs-cl
27 Apr 2026
Research

Mixed Membership sub-Gaussian Models

DGX agent

arXiv:2604.22633v1 Announce Type: cross Abstract: The Gaussian mixture model is widely used in unsupervised learning, owing to its simplicity and interpretability. However, a fundamental limitation of

researcharxiv-cs-lg
27 Apr 2026
Model Releases

On Benchmark Hacking in ML Contests: Modeling, Insights and Design

DGX agent

arXiv:2604.22230v1 Announce Type: cross Abstract: Benchmark hacking refers to tuning a machine learning model to score highly on certain evaluation criteria without improving true generalization or fa

model-releasesarxiv-cs-lg
27 Apr 2026
Model Releases

Towards Adaptive Continual Model Merging via Manifold-Aware Expert Evolution

DGX agent

arXiv:2604.22464v1 Announce Type: new Abstract: Continual Model Merging (CMM) sequentially integrates task-specific models into a unified architecture without intensive retraining. However, existing C

model-releasesarxiv-cs-lg
27 Apr 2026
Research

Algebraic Language Models for Inverse Design of Metamaterials via Diffusion Transformers

DGX agent

arXiv:2507.15753v2 Announce Type: replace-cross Abstract: Generative machine learning models have revolutionized material discovery by capturing complex structure-property relationships, yet extending

researcharxiv-cs-ai
24 Apr 2026
Research

Context Unrolling in Omni Models

DGX agent

arXiv:2604.21921v1 Announce Type: new Abstract: We present Omni, a unified multimodal model natively trained on diverse modalities, including text, images, videos, 3D geometry, and hidden representati

researcharxiv-cs-cv
24 Apr 2026
Model Releases

DenoiseRank: Learning to Rank by Diffusion Models

DGX agent

arXiv:2604.20852v1 Announce Type: cross Abstract: Learning to rank (LTR) is one of the core tasks in Machine Learning. Traditional LTR models have made great progress, but nearly all of them are imple

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Efficient Multi-Source Knowledge Transfer by Model Merging

DGX agent

arXiv:2508.19353v2 Announce Type: replace-cross Abstract: While transfer learning is an effective strategy, it often overlooks the opportunity to leverage knowledge from numerous available models onli

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

HyperFM: An Efficient Hyperspectral Foundation Model with Spectral Grouping

DGX agent

arXiv:2604.21127v1 Announce Type: new Abstract: The NASA PACE mission provides unprecedented hyperspectral observations of ocean color, aerosols, and clouds, offering new insights into how these compo

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

Latent Denoising Improves Visual Alignment in Large Multimodal Models

DGX agent

arXiv:2604.21343v1 Announce Type: new Abstract: Large Multimodal Models (LMMs) such as LLaVA are typically trained with an autoregressive language modeling objective, providing only indirect supervisi

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

mGRADE: Minimal Recurrent Gating Meets Delay Convolutions for Lightweight Sequence Modeling

DGX agent

arXiv:2507.01829v2 Announce Type: replace-cross Abstract: Multi-timescale sequence modeling relies on capturing both local fast dynamics and global slow context; yet, maintaining these capabilities un

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Open-H-Embodiment: A Large-Scale Dataset for Enabling Foundation Models in Medical Robotics

DGX agent

arXiv:2604.21017v1 Announce Type: cross Abstract: Autonomous medical robots hold promise to improve patient outcomes, reduce provider workload, democratize access to care, and enable superhuman precis

model-releasesarxiv-cs-ai
24 Apr 2026
← Previous
1…5253545556…1021
Next →