AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,522 results
Safety

An Information-Geometric Framework for Stability Analysis of Large Language Models under Entropic Stress

DGX agent

arXiv:2604.24076v1 Announce Type: new Abstract: As large language models (LLMs) are increasingly deployed in high-stakes and operational settings, evaluation strategies based solely on aggregate accur

safetyarxiv-cs-ai
28 Apr 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

AutoPyVerifier: Learning Compact Executable Verifiers for Large Language Model Outputs

DGX agent

arXiv:2604.22937v1 Announce Type: new Abstract: Verification is becoming central to both reinforcement-learning-based training and inference-time control of large language models (LLMs). Yet current v

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

Benchmarking and Mitigating Sycophancy in Medical Vision Language Models

DGX agent

arXiv:2509.21979v4 Announce Type: replace-cross Abstract: Visual language models (VLMs) have the potential to transform medical workflows. However, the deployment is limited by sycophancy. Despite thi

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Can Multimodal Large Language Models Truly Understand Small Objects?

DGX agent

arXiv:2604.22884v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have shown promising potential in diverse understanding tasks, e.g., image and video analysis, math and physi

model-releasesarxiv-cs-ai
28 Apr 2026
Research

Characterizing Vision-Language-Action Models across XPUs: Constraints and Acceleration for On-Robot Deployment

DGX agent

arXiv:2604.24447v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models are promising for generalist robot control, but on-robot deployment is bottlenecked by real-time inference under t

researcharxiv-cs-ai
28 Apr 2026
Research

Contextual Linear Activation Steering of Language Models

DGX agent

arXiv:2604.24693v1 Announce Type: new Abstract: Linear activation steering is a powerful approach for eliciting the capabilities of large language models and specializing their behavior using limited

researcharxiv-cs-cl
28 Apr 2026
Safety

Discovering Failure Modes in Vision-Language Models using RL

DGX agent

arXiv:2604.04733v2 Announce Type: replace-cross Abstract: Vision-language Models (VLMs), despite achieving strong performance on multimodal benchmarks, often misinterpret straightforward visual concep

safetyarxiv-cs-ai
28 Apr 2026
Model Releases

DO-Bench: An Attributable Benchmark for Diagnosing Object Hallucination in Vision-Language Models

DGX agent

arXiv:2604.22822v1 Announce Type: cross Abstract: Object level hallucination remains a central reliability challenge for vision language models (VLMs), particularly in binary object existence verifica

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

EAGLE: Expert-Augmented Attention Guidance for Tuning-Free Industrial Anomaly Detection in Multimodal Large Language Models

DGX agent

arXiv:2602.17419v3 Announce Type: replace Abstract: Multimodal large language models (MLLMs) can enrich industrial anomaly detection with semantic descriptions and anomaly reasoning, but they still la

model-releasesarxiv-cs-cv
28 Apr 2026
Applications

FedRef: Bayesian Fine-Tuning using a Reference Model to Mitigate Catastrophic Forgetting for Heterogeneous Federated Learning

DGX agent

arXiv:2506.23210v5 Announce Type: replace-cross Abstract: Federated learning (FL) enables collaborative model training across distributed clients while preserving data privacy. However, data and syste

applicationsarxiv-cs-ai
28 Apr 2026
Tutorials

Fine-tuning vs. In-context Learning in Large Language Models: A Formal Language Learning Perspective

DGX agent

arXiv:2604.23267v1 Announce Type: new Abstract: Large language models (LLMs) operate in two fundamental learning modes - fine-tuning (FT) and in-context learning (ICL) - raising key questions about wh

tutorialsarxiv-cs-cl
28 Apr 2026
Research

HeadRouter: Dynamic Head-Weight Routing for Task-Adaptive Audio Token Pruning in Large Audio Language Models

DGX agent

arXiv:2604.23717v1 Announce Type: cross Abstract: Recent large audio language models (LALMs) demonstrate remarkable capabilities in processing extended multi-modal sequences, yet incur high inference

researcharxiv-cs-cl
28 Apr 2026
Applications

HeiSD: Hybrid Speculative Decoding for Embodied Vision-Language-Action Models with Kinematic Awareness

DGX agent

arXiv:2603.17573v2 Announce Type: replace-cross Abstract: Vision-Language-Action (VLA) Models have become the mainstream solution for robot control, but suffer from slow inference speeds. Speculative

applicationsarxiv-cs-lg
28 Apr 2026
Research

Instruction-Free Tuning of Large Vision Language Models for Medical Instruction Following

DGX agent

arXiv:2603.19482v2 Announce Type: replace Abstract: Large vision language models (LVLMs) have demonstrated impressive performance across a wide range of tasks. These capabilities largely stem from vis

researcharxiv-cs-cv
28 Apr 2026
Tutorials

Inverting Foundation Models of Brain Function with Simulation-Based Inference

DGX agent

arXiv:2604.23865v1 Announce Type: cross Abstract: Foundation models of brain activity promise a new frontier for in silico neuroscience by emulating neural responses to complex stimuli across tasks an

tutorialsarxiv-cs-ai
28 Apr 2026
Model Releases

Large Language Models as Virtual Survey Respondents: Evaluating Sociodemographic Response Generation

DGX agent

arXiv:2509.06337v2 Announce Type: replace Abstract: Questionnaire-based surveys are foundational to social science research and public policymaking, yet traditional survey methods remain costly, time-

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Less Is More: Engineering Challenges of On-Device Small Language Model Integration in a Mobile Application

DGX agent

arXiv:2604.24636v1 Announce Type: cross Abstract: On-device Small Language Models (SLMs) promise fully offline, private AI experiences for mobile users (no cloud dependency, no data leaving the device

model-releasesarxiv-cs-ai
28 Apr 2026
Research

Live Knowledge Tracing: Real-Time Adaptation using Tabular Foundation Models

DGX agent

arXiv:2602.06542v3 Announce Type: replace Abstract: Deep knowledge tracing models have achieved significant breakthroughs in modeling student learning trajectories. However, these architectures requir

researcharxiv-cs-lg
28 Apr 2026
Model Releases

Machine Learning and Deep Learning Models for Short Term Electricity Price Forecasting in Australia's National Electricity Market

DGX agent

arXiv:2604.23908v1 Announce Type: new Abstract: Short term electricity price forecast is essential in competitive power markets, yet electricity price series exhibit high volatility, irregularity, and

model-releasesarxiv-cs-lg
28 Apr 2026
Research

Machine learning models for estimating counterfactuals in a single-arm inflammatory bowel disease study

DGX agent

arXiv:2604.23465v1 Announce Type: new Abstract: Single-arm trials accelerate study timelines by reducing the number of patients that must be recruited for a concurrent control group. However, these de

researcharxiv-cs-lg
28 Apr 2026
Model Releases

Nemotron-3-Nano-Omni-30B-A3B-Reasoning, New model?

DGX agent

NVIDIA Nemotron 3 Nano Omni is a multimodal large language model that unifies video, audio, image, and text understanding for enterprise Q&A, summarization, transcription, and document intelligence, w

model-releasesr-localllama
28 Apr 2026
Model Releases

Nvidia launches Nemotron 3 Nano Omni, an open multimodal model with a 30B-A3B hybrid MoE architecture; the Nemotron 3 family saw 50M+ downloads in the past year (Kyt Dotson/SiliconANGLE)

DGX agent

Kyt Dotson / SiliconANGLE: Nvidia launches Nemotron 3 Nano Omni, an open multimodal model with a 30B-A3B hybrid MoE architecture; the Nemotron 3 family saw 50M+ downloads in the past year — Nvidia Cor

model-releasestechmeme
28 Apr 2026
Research

NVILA: Efficient Frontier Visual Language Models

DGX agent

arXiv:2412.04468v3 Announce Type: replace Abstract: Visual language models (VLMs) have made significant advances in accuracy in recent years. However, their efficiency has received much less attention

researcharxiv-cs-cv
28 Apr 2026
Model Releases

OpenAI models, Codex, and Managed Agents come to AWS

DGX agent

OpenAI announced the availability of its models, including Codex, and managed agent capabilities on Amazon Web Services (AWS) infrastructure. This integration enables AWS customers to access OpenAI's

model-releasesopenai
28 Apr 2026
Model Releases

PDF-WuKong: A Large Multimodal Model for Efficient Long PDF Reading with End-to-End Sparse Sampling

DGX agent

arXiv:2410.05970v3 Announce Type: replace-cross Abstract: Multimodal document understanding is a challenging task to process and comprehend large amounts of textual and visual information. Recent adva

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Seeing Is No Longer Believing: Frontier Image Generation Models, Synthetic Visual Evidence, and Real-World Risk

DGX agent

arXiv:2604.24197v1 Announce Type: cross Abstract: Frontier image generation has moved from artistic synthesis toward synthetic visual evidence. Systems such as GPT Image 2, Nano Banana Pro, Nano Banan

model-releasesarxiv-cs-ai
28 Apr 2026
Research

Small Language Model Helps Resolve Semantic Ambiguity of LLM Prompt

DGX agent

arXiv:2604.23263v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly utilized in various complex reasoning tasks due to their excellent instruction following capability. How

researcharxiv-cs-ai
28 Apr 2026
Model Releases

Speech Enhancement Based on Drifting Models

DGX agent

arXiv:2604.24199v1 Announce Type: cross Abstract: We propose Speech Enhancement based on Drifting Models (DriftSE), a novel generative framework that formulates denoising as an equilibrium problem. Ra

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Think-at-Hard: Selective Latent Iterations to Improve Reasoning Language Models

DGX agent

arXiv:2511.08577v2 Announce Type: replace-cross Abstract: Improving reasoning abilities of Large Language Models (LLMs), especially under parameter constraints, is crucial for real-world applications.

model-releasesarxiv-cs-ai
28 Apr 2026
Agents

Today we’re shipping Laguna M.1 and Laguna XS.2 – our first public models. We’re also shipping our agent harness and a preview product exper…

DGX agent

Today we’re shipping Laguna M.1 and Laguna XS.2 – our first public models. We’re also shipping our agent harness and a preview product experience. Both models were trained from scratch on our own stac

agentsclem-delangue--x
28 Apr 2026
Local Ai

Unified Multi-Foundation-Model Slide Representation for Pan-Cancer Recognition and Text-Guided Tumor Localization

DGX agent

arXiv:2604.22846v1 Announce Type: new Abstract: The expanding ecosystem of pathology foundation models has produced powerful but fragmented tile-level representations, limiting their use in clinical t

local-aiarxiv-cs-cv
28 Apr 2026
Model Releases

Can Large Language Models Adequately Perform Symbolic Reasoning Over Time Series?

DGX agent

arXiv:2508.03963v4 Announce Type: replace Abstract: Uncovering hidden symbolic laws from time series data, as an aspiration dating back to Kepler's discovery of planetary motion, remains a core challe

model-releasesarxiv-cs-ai
27 Apr 2026
Model Releases

FILTR: Extracting Topological Features from Pretrained 3D Models

DGX agent

arXiv:2604.22334v1 Announce Type: new Abstract: Recent advances in pretraining 3D point cloud encoders (e.g., Point-BERT, Point-MAE) have produced powerful models, whose abilities are typically evalua

model-releasesarxiv-cs-cv
27 Apr 2026
Local Ai

Fine-Grained Analysis of Shared Syntactic Mechanisms in Language Models

DGX agent

arXiv:2604.22166v1 Announce Type: new Abstract: While language models demonstrate sophisticated syntactic capabilities, the extent to which their internal mechanisms align with cross-constructional pr

local-aiarxiv-cs-cl
27 Apr 2026
Model Releases

From Interpretability to Performance: Optimizing Retrieval Heads for Long-Context Language Models

DGX agent

arXiv:2601.11020v3 Announce Type: replace Abstract: Advances in mechanistic interpretability have identified special attention heads, known as retrieval heads, that are responsible for retrieving info

model-releasesarxiv-cs-cl
27 Apr 2026
Model Releases

going to buy 2 rtx 6000 just because of the capabilities of local models becoming great! no more outsourcing of research to the api

DGX agent

going to buy 2 rtx 6000 just because of the capabilities of local models becoming great! no more outsourcing of research to the api Xiaomi MiMo-V2.5 is now officially open-sourced! MIT License, suppor

model-releasesclem-delangue--x
27 Apr 2026
Model Releases

Graph-to-Vision: Multi-graph Understanding and Reasoning using Vision-Language Models

DGX agent

arXiv:2503.21435v3 Announce Type: replace Abstract: Recent advances in Vision-Language Models (VLMs) have shown promising capabilities in interpreting visualized graph data, offering a new perspective

model-releasesarxiv-cs-ai
27 Apr 2026
Local Ai

MambaCSP: Hybrid-Attention State Space Models for Hardware-Efficient Channel State Prediction

DGX agent

arXiv:2604.21957v1 Announce Type: cross Abstract: Recent works have demonstrated that attention-based transformer and large language model (LLM) architectures can achieve strong channel state predicti

local-aiarxiv-cs-ai
27 Apr 2026
Applications

Mochi: Aligning Pre-training and Inference for Efficient Graph Foundation Models via Meta-Learning

DGX agent

arXiv:2604.22031v1 Announce Type: cross Abstract: We propose Mochi, a Graph Foundation Model that addresses task unification and training efficiency by adopting a meta-learning based training framewor

applicationsarxiv-cs-ai
27 Apr 2026
Model Releases

MTT-Bench: Predicting Social Dominance in Mice via Multimodal Large Language Models

DGX agent

arXiv:2604.22492v1 Announce Type: cross Abstract: Understanding social dominance in animal behavior is critical for neuroscience and behavioral studies. In this work, we explore the capability of Mult

model-releasesarxiv-cs-cv
27 Apr 2026
Hardware

Multimodal Neural Operators for Real-Time Biomechanical Modelling of Traumatic Brain Injury

DGX agent

arXiv:2510.03248v3 Announce Type: replace-cross Abstract: Background: Traumatic brain injury modeling requires integrating volumetric neuroimaging, demographic parameters, and acquisition metadata. Fi

hardwarearxiv-cs-ai
27 Apr 2026
Applications

Nuclear Diffusion Models for Low-Rank Background Suppression in Videos

DGX agent

arXiv:2509.20886v2 Announce Type: replace Abstract: Video sequences often contain structured noise and background artifacts that obscure dynamic content, posing challenges for accurate analysis and re

applicationsarxiv-cs-cv
27 Apr 2026
Model Releases

Relaxation-Informed Training of Neural Network Surrogate Models

DGX agent

arXiv:2604.22746v1 Announce Type: cross Abstract: ReLU neural networks trained as surrogate models can be embedded exactly in mixed-integer linear programs (MILPs), enabling global optimization over t

model-releasesarxiv-cs-lg
27 Apr 2026
Model Releases

Score-based Membership Inference on Diffusion Models

DGX agent

arXiv:2509.25003v2 Announce Type: replace-cross Abstract: Membership inference attacks (MIAs) against Diffusion Models (DMs) raise pressing privacy concerns by revealing whether a sample was part of t

model-releasesarxiv-cs-cv
27 Apr 2026
Hardware

SpikingBrain2.0: Brain-Inspired Foundation Models for Efficient Long-Context and Cross-Platform Inference

DGX agent

arXiv:2604.22575v1 Announce Type: new Abstract: Scaling context length is reshaping large-model development, yet full-attention Transformers suffer from prohibitive computation and inference bottlenec

hardwarearxiv-cs-lg
27 Apr 2026
Research

Spontaneous Persuasion: An Audit of Model Persuasiveness in Everyday Conversations

DGX agent

arXiv:2604.22109v1 Announce Type: cross Abstract: Large language models (LLMs) possess strong persuasive capabilities that outperform humans in head-to-head comparisons. Users report consulting LLMs t

researcharxiv-cs-ai
27 Apr 2026
Industry

Top 3 trending models of the week on HF: @deepseek_ai @OpenAI & @Alibaba_Qwen!

DGX agent

This post highlights the three most popular models on Hugging Face during a given week, featuring DeepSeek AI, OpenAI, and Alibaba's Qwen models. The post was shared by Clem Delangue, CEO of Hugging F

industryclem-delangue--x
27 Apr 2026
Safety

TTS-PRISM: A Perceptual Reasoning and Interpretable Speech Model for Fine-Grained Diagnosis

DGX agent

arXiv:2604.22225v1 Announce Type: new Abstract: While generative text-to-speech (TTS) models approach human-level quality, monolithic metrics fail to diagnose fine-grained acoustic artifacts or explai

safetyarxiv-cs-cl
27 Apr 2026
← Previous
1…115116117118119…1261
Next →