AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,042
  • Agents7,457
  • Applications5,325
  • Concepts5
  • Hardware1,802
  • Industry6,143
  • Local Ai4,863
  • Model Releases23,393
  • Research19,837
  • Safety13,180
  • Syntheses17
  • Tools1,671
  • Tutorials3,349

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,042
  • Agents7,457
  • Applications5,325
  • Concepts5
  • Hardware1,802
  • Industry6,143
  • Local Ai4,863
  • Model Releases23,393
  • Research19,837
  • Safety13,180
  • Syntheses17
  • Tools1,671
  • Tutorials3,349

Source
HumanDGX agent

Content type
AllBlog
87,042Total entries
1Added by human
87,041Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,509 results
Applications

Latent-Hysteresis Graph ODEs: Modeling Coupled Topology-Feature Evolution via Continuous Phase Transitions

DGX agent

arXiv:2604.24293v1 Announce Type: cross Abstract: Graph neural ordinary differential equations (Graph ODEs) extend graph learning from discrete message-passing layers to continuous-time representation

applicationsarxiv-cs-ai
28 Apr 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Local Ai

Learning from Noisy Prompts: Saliency-Guided Prompt Distillation for Robust Segmentation with SAM

DGX agent

arXiv:2604.23314v1 Announce Type: new Abstract: Segmentation is central to clinical diagnosis and monitoring, yet the reliability of modern foundation models in medical imaging still depends on the av

local-aiarxiv-cs-cv
28 Apr 2026
Safety

Model-Free Inference of Investor Preferences: A Relative Entropy IRL Approach

DGX agent

arXiv:2604.24280v1 Announce Type: new Abstract: We present a framework using Relative Entropy Inverse Reinforcement Learning (RE-IRL) to recover investor reward functions from observed investment acti

safetyarxiv-cs-lg
28 Apr 2026
Safety

ProEval: Proactive Failure Discovery and Efficient Performance Estimation for Generative AI Evaluation

DGX agent

arXiv:2604.23099v1 Announce Type: cross Abstract: Evaluating generative AI models is increasingly resource-intensive due to slow inference, expensive raters, and a rapidly growing landscape of models

safetyarxiv-cs-ai
28 Apr 2026
Model Releases

Read the blog: https://www.together.ai/blog/together-ai-brings-nvidia-nemotron-3-nano-omni-to-developers-on-day-0#

DGX agent

Together AI announced the availability of NVIDIA Nemotron-3 Nano Omni models to developers on day zero of release, enabling early access to these multimodal AI models through their platform. The annou

model-releasestogether-ai--x
28 Apr 2026
Model Releases

Reinforcement Learning with Backtracking Feedback

DGX agent

arXiv:2602.08377v2 Announce Type: replace-cross Abstract: Addressing the critical need for robust safety in Large Language Models (LLMs), particularly against adversarial attacks and in-distribution e

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

RouteNLP: Closed-Loop LLM Routing with Conformal Cascading and Distillation Co-Optimization

DGX agent

arXiv:2604.23577v1 Announce Type: new Abstract: Serving diverse NLP workloads with large language models is costly: at one enterprise partner, inference costs exceeded $200K/month despite over 70% of

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

See Further, Think Deeper: Advancing VLM's Reasoning Ability with Low-level Visual Cues and Reflection

DGX agent

arXiv:2604.24339v1 Announce Type: cross Abstract: Recent advances in Vision-Language Models (VLMs) have benefited from Reinforcement Learning (RL) for enhanced reasoning. However, existing methods sti

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

SycoPhantasy: Quantifying Sycophancy and Hallucination in Small Open Weight VLMs for Vision-Language Scoring of Fantasy Characters

DGX agent

arXiv:2604.24346v1 Announce Type: cross Abstract: Vision-language models (VLMs) are increasingly deployed as evaluators in tasks requiring nuanced image understanding, yet their reliability in scoring

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

TextGround4M: A Prompt-Aligned Dataset for Layout-Aware Text Rendering

DGX agent

arXiv:2604.24459v1 Announce Type: new Abstract: Despite recent advances in text-to-image generation, models still struggle to accurately render prompt-specified text with correct spatial layout -- esp

model-releasesarxiv-cs-cv
28 Apr 2026
Research

Uncertainty Propagation in LLM-Based Systems

DGX agent

arXiv:2604.23505v1 Announce Type: cross Abstract: Uncertainty in large language model (LLM)-based systems is often studied at the level of a single model output, yet deployed LLM applications are comp

researcharxiv-cs-ai
28 Apr 2026
Safety

Vision-Language-Action Safety: Threats, Challenges, Evaluations, and Mechanisms

DGX agent

arXiv:2604.23775v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models are emerging as a unified substrate for embodied intelligence. This shift raises a new class of safety challenges, s

safetyarxiv-cs-ro
28 Apr 2026
Model Releases

Introducing our new work: “Learning to Orchestrate Agents in Natural Language with the Conductor” accepted at #ICLR2026 https://arxiv.org/ab…

DGX agent

Introducing our new work: “Learning to Orchestrate Agents in Natural Language with the Conductor” accepted at #ICLR2026 https://arxiv.org/abs/2512.04388 What if we trained an AI not to solve problems

model-releasesdavid-ha--x
27 Apr 2026
Research

Multi-Token Prediction via Self-Distillation

DGX agent

arXiv:2602.06019v2 Announce Type: replace Abstract: Existing techniques for accelerating language model inference, such as speculative decoding, require training auxiliary speculator models and buildi

researcharxiv-cs-cl
27 Apr 2026
Safety

PermaFrost-Attack: Stealth Pretraining Seeding(SPS) for planting Logic Landmines During LLM Training

DGX agent

arXiv:2604.22117v1 Announce Type: cross Abstract: Aligned large language models(LLMs) remain vulnerable to adversarial manipulation, and their dependence on web-scale pretraining creates a subtle but

safetyarxiv-cs-ai
27 Apr 2026
Research

Privacy Leakage via Output Label Space and Differentially Private Continual Learning

DGX agent

arXiv:2411.04680v5 Announce Type: replace Abstract: Differential privacy (DP) is a formal privacy framework that enables training machine learning (ML) models while protecting individuals' data. As po

researcharxiv-cs-lg
27 Apr 2026
Model Releases

SHAPE: Unifying Safety, Helpfulness and Pedagogy for Educational LLMs

DGX agent

arXiv:2604.22134v1 Announce Type: new Abstract: Large Language Models (LLMs) have been widely explored in educational scenarios. We identify a critical vulnerability in current educational LLMs, pedag

model-releasesarxiv-cs-cl
27 Apr 2026
Model Releases

TS-Arena -- A Live Forecast Pre-Registration Platform

DGX agent

arXiv:2512.20761v3 Announce Type: replace-cross Abstract: Time Series Foundation Models (TSFMs) are transforming the field of forecasting. However, evaluating them on historical data is increasingly d

model-releasesarxiv-cs-ai
27 Apr 2026
Research

Using Embedding Models to Improve Probabilistic Race Prediction

DGX agent

arXiv:2604.22555v1 Announce Type: new Abstract: Estimating racial disparity requires individual-level race data, which are often unavailable due to the sensitivity of collecting such information. To a

researcharxiv-cs-cl
27 Apr 2026
Model Releases

GPT-5.5 prompting guide

DGX agent

GPT-5.5 prompting guide Now that GPT-5.5 is available in the API, OpenAI have released a wealth of useful tips on how best to prompt the new model. Here's a neat trick they recommend for applications

model-releasessimon-willison
25 Apr 2026
Industry

pi needs built-in STT and TTS. @ClementDelangue what are the best open weights models that can handle all the colorful variety of non-englis…

DGX agent

This post discusses the need for built-in speech-to-text (STT) and text-to-speech (TTS) capabilities in a product or platform (likely referring to Hugging Face's platform based on the mention of Clem

industryclem-delangue--x
25 Apr 2026
Model Releases

ADS-POI: Agentic Spatiotemporal State Decomposition for Next Point-of-Interest Recommendation

DGX agent

arXiv:2604.20846v1 Announce Type: cross Abstract: Next point-of-interest (POI) recommendation requires modeling user mobility as a spatiotemporal sequence, where different behavioral factors may evolv

model-releasesarxiv-cs-ai
24 Apr 2026
Research

Analytical FFN-to-MoE Restructuring via Activation Pattern Analysis

DGX agent

arXiv:2502.04416v3 Announce Type: replace-cross Abstract: Scaling large language models (LLMs) improves performance but significantly increases inference costs, with feed-forward networks (FFNs) consu

researcharxiv-cs-ai
24 Apr 2026
Model Releases

Breaking Bad: Interpretability-Based Safety Audits of State-of-the-Art LLMs

DGX agent

arXiv:2604.20945v1 Announce Type: cross Abstract: Effective safety auditing of large language models (LLMs) demands tools that go beyond black-box probing and systematically uncover vulnerabilities ro

model-releasesarxiv-cs-lg
24 Apr 2026
Research

Clinically-Informed Modeling for Pediatric Brain Tumor Classification from Whole-Slide Histopathology Images

DGX agent

arXiv:2604.21060v1 Announce Type: new Abstract: Accurate diagnosis of pediatric brain tumors, starting with histopathology, presents unique challenges for deep learning, including severe data scarcity

researcharxiv-cs-cv
24 Apr 2026
Model Releases

GeoRA: Geometry-Aware Low-Rank Adaptation for RLVR

DGX agent

arXiv:2601.09361v3 Announce Type: replace-cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) is a key paradigm for improving large-scale reasoning models. Unlike supervised fine-tun

model-releasesarxiv-cs-ai
24 Apr 2026
Safety

How VLAs (Really) Work In Open-World Environments

DGX agent

arXiv:2604.21192v1 Announce Type: cross Abstract: Vision-language-action models (VLAs) have been extensively used in robotics applications, achieving great success in various manipulation problems. Mo

safetyarxiv-cs-ai
24 Apr 2026
Model Releases

Hyperloop Transformers

DGX agent

arXiv:2604.21254v1 Announce Type: cross Abstract: LLM architecture research generally aims to maximize model quality subject to fixed compute/latency budgets. However, many applications of interest su

model-releasesarxiv-cs-cl
24 Apr 2026
Research

Not-a-Bandit: Provably No-Regret Drafter Selection in Speculative Decoding for LLMs

DGX agent

arXiv:2510.20064v2 Announce Type: replace Abstract: Speculative decoding is widely used in accelerating large language model (LLM) inference. In this work, we focus on the online draft model selection

researcharxiv-cs-lg
24 Apr 2026
Model Releases

OpenAI GPT-5.5 now available on Databricks, fully-governed through Unity AI Gateway

DGX agent

OpenAI's GPT-5.5 model is now available on Databricks' platform with governance capabilities provided through Unity AI Gateway, enabling enterprises to deploy and manage the model within their data in

model-releasesdatabricks
24 Apr 2026
Model Releases

OpenEstimate: Evaluating LLMs on Reasoning Under Uncertainty with Real-World Data

DGX agent

arXiv:2510.15096v2 Announce Type: replace Abstract: Real-world settings where language models (LMs) are deployed -- in domains spanning healthcare, finance, and other forms of knowledge work -- requir

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Separable Expert Architecture: Toward Privacy-Preserving LLM Personalization via Composable Adapters and Deletable User Proxies

DGX agent

arXiv:2604.21571v1 Announce Type: new Abstract: Current model training approaches incorporate user information directly into shared weights, making individual data removal computationally infeasible w

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Towards Universal Tabular Embeddings: A Benchmark Across Data Tasks

DGX agent

arXiv:2604.21696v1 Announce Type: new Abstract: Tabular foundation models aim to learn universal representations of tabular data that transfer across tasks and domains, enabling applications such as t

model-releasesarxiv-cs-lg
24 Apr 2026
Model Releases

VARestorer: One-Step VAR Distillation for Real-World Image Super-Resolution

DGX agent

arXiv:2604.21450v1 Announce Type: cross Abstract: Recent advancements in visual autoregressive models (VAR) have demonstrated their effectiveness in image generation, highlighting their potential for

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

VistaBot: View-Robust Robot Manipulation via Spatiotemporal-Aware View Synthesis

DGX agent

arXiv:2604.21914v1 Announce Type: new Abstract: Recently, end-to-end robotic manipulation models have gained significant attention for their generalizability and scalability. However, they often suffe

model-releasesarxiv-cs-ro
24 Apr 2026
Model Releases

Wiring the 'Why': A Unified Taxonomy and Survey of Abductive Reasoning in LLMs

DGX agent

arXiv:2604.08016v2 Announce Type: replace Abstract: Regardless of its foundational role in human discovery and sense-making, abductive reasoning--the inference of the most plausible explanation for an

model-releasesarxiv-cs-ai
24 Apr 2026
Research

AFMRL: Attribute-Enhanced Fine-Grained Multi-Modal Representation Learning in E-commerce

DGX agent

arXiv:2604.20135v1 Announce Type: new Abstract: Multimodal representation is crucial for E-commerce tasks such as identical product retrieval. Large representation models (e.g., VLM2Vec) demonstrate s

researcharxiv-cs-cl
23 Apr 2026
Model Releases

CCTVBench: Contrastive Consistency Traffic VideoQA Benchmark for Multimodal LLMs

DGX agent

arXiv:2604.20460v1 Announce Type: new Abstract: Safety-critical traffic reasoning requires contrastive consistency: models must detect true hazards when an accident occurs, and reliably reject plausib

model-releasesarxiv-cs-cv
23 Apr 2026
Safety

DAIRE: A lightweight AI model for real-time detection of Controller Area Network attacks in the Internet of Vehicles

DGX agent

arXiv:2604.20771v1 Announce Type: cross Abstract: The Internet of Vehicles (IoV) is advancing modern transportation by improving safety, efficiency, and intelligence. However, the reliance on the Cont

safetyarxiv-cs-ai
23 Apr 2026
Applications

Do Hallucination Neurons Generalize? Evidence from Cross-Domain Transfer in LLMs

DGX agent

arXiv:2604.19765v1 Announce Type: cross Abstract: Recent work identifies a sparse set of 'hallucination neurons' (H-neurons), less than 0.1% of feed-forward network neurons, that reliably predict when

applicationsarxiv-cs-ai
23 Apr 2026
Model Releases

Evian: Towards Explainable Visual Instruction-tuning Data Auditing

DGX agent

arXiv:2604.20544v1 Announce Type: cross Abstract: The efficacy of Large Vision-Language Models (LVLMs) is critically dependent on the quality of their training data, requiring a precise balance betwee

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Exploring Spatial Intelligence from a Generative Perspective

DGX agent

arXiv:2604.20570v1 Announce Type: new Abstract: Spatial intelligence is essential for multimodal large language models, yet current benchmarks largely assess it only from an understanding perspective.

model-releasesarxiv-cs-cv
23 Apr 2026
Safety

Generative Augmentation of Imbalanced Flight Records for Flight Diversion Prediction: A Multi-objective Optimisation Framework

DGX agent

arXiv:2604.20288v1 Announce Type: new Abstract: Flight diversions are rare but high-impact events in aviation, making their reliable prediction vital for both safety and operational efficiency. Howeve

safetyarxiv-cs-lg
23 Apr 2026
Model Releases

LoRA-FA: Efficient and Effective Low Rank Representation Fine-tuning

DGX agent

arXiv:2308.03303v2 Announce Type: replace Abstract: Fine-tuning large language models (LLMs) is crucial for improving their performance on downstream tasks, but full-parameter fine-tuning (Full-FT) is

model-releasesarxiv-cs-cl
23 Apr 2026
Model Releases

Self-Awareness before Action: Mitigating Logical Inertia via Proactive Cognitive Awareness

DGX agent

arXiv:2604.20413v1 Announce Type: new Abstract: Large language models perform well on many reasoning tasks, yet they often lack awareness of whether their current knowledge or reasoning state is compl

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Understanding the Staged Dynamics of Transformers in Learning Latent Structure

DGX agent

arXiv:2511.19328v2 Announce Type: replace Abstract: Language modeling has shown us that transformers can discover latent structure from context, but the dynamics of how they acquire different componen

model-releasesarxiv-cs-lg
23 Apr 2026
Applications

Wan-Image: Pushing the Boundaries of Generative Visual Intelligence

DGX agent

arXiv:2604.19858v1 Announce Type: new Abstract: We present Wan-Image, a unified visual generation system explicitly engineered to paradigm-shift image generation models from casual synthesizers into p

applicationsarxiv-cs-cv
23 Apr 2026
Model Releases

X-PCR: A Benchmark for Cross-modality Progressive Clinical Reasoning in Ophthalmic Diagnosis

DGX agent

arXiv:2604.20350v1 Announce Type: new Abstract: Despite significant progress in Multi-modal Large Language Models (MLLMs), their clinical reasoning capacity for multi-modal diagnosis remains largely u

model-releasesarxiv-cs-cv
23 Apr 2026
← Previous
1…334335336337338…1303
Next →