AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,998
  • Agents7,605
  • Applications5,444
  • Concepts5
  • Hardware1,850
  • Industry6,179
  • Local Ai4,964
  • Model Releases24,131
  • Research20,257
  • Safety13,456
  • Syntheses17
  • Tools1,677
  • Tutorials3,413

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,998
  • Agents7,605
  • Applications5,444
  • Concepts5
  • Hardware1,850
  • Industry6,179
  • Local Ai4,964
  • Model Releases24,131
  • Research20,257
  • Safety13,456
  • Syntheses17
  • Tools1,677
  • Tutorials3,413

Source
HumanDGX agent

Content type
88,998Total entries
1Added by human
88,997Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
64,133 results
Research

From Syntax to Emotion: A Mechanistic Analysis of Emotion Inference in LLMs

DGX agent

arXiv:2604.25866v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used in emotionally sensitive human-AI applications, yet little is known about how emotion recognition is

researcharxiv-cs-cl
29 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

MGSM-Pro: A Simple Strategy for Robust Multilingual Mathematical Reasoning Evaluation

DGX agent

arXiv:2601.21225v2 Announce Type: replace Abstract: Large language models have made substantial progress in mathematical reasoning. However, benchmark development for multilingual evaluation has lagge

model-releasesarxiv-cs-cl
29 Apr 2026
Research

Mutual Forcing: Dual-Mode Self-Evolution for Fast Autoregressive Audio-Video Character Generation

DGX agent

arXiv:2604.25819v1 Announce Type: new Abstract: In this work, we propose Mutual Forcing, a framework for fast autoregressive audio-video generation with long-horizon audio-video synchronization. Our a

researcharxiv-cs-cv
29 Apr 2026
Research

Novel 3D Binary Indexed Tree for Volume Computation of 3D Reconstructed Models from Volumetric Data

DGX agent

arXiv:2412.10441v2 Announce Type: replace-cross Abstract: In the burgeoning field of medical imaging, precise computation of 3D volume holds a significant importance for subsequent qualitative analysi

researcharxiv-cs-cv
29 Apr 2026
Model Releases

OMHBench: Benchmarking Balanced and Grounded Omni-Modal Multi-Hop Reasoning

DGX agent

arXiv:2508.16198v3 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) have increasingly supported omni-modal processing across text, vision, and speech. However, existing evalua

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

One Perturbation, Two Failure Modes: Probing VLM Safety via Embedding-Guided Typographic Perturbations

DGX agent

arXiv:2604.25102v1 Announce Type: new Abstract: Typographic prompt injection exploits vision language models' (VLMs) ability to read text rendered in images, posing a growing threat as VLMs power auto

model-releasesarxiv-cs-cv
29 Apr 2026
Research

Quantum-Inspired Robust and Scalable SAR Object Classification

DGX agent

arXiv:2604.25755v1 Announce Type: cross Abstract: SAR image classification naturally has to deal with huge noise and a high dynamic range particularly requiring robust classification models. Additiona

researcharxiv-cs-cv
29 Apr 2026
Safety

A Self-Supervised Framework for Space Object Behaviour Characterisation

DGX agent

arXiv:2504.06176v3 Announce Type: replace-cross Abstract: Foundation Models, which leverage large neural networks pre-trained on unlabelled data before fine-tuning for specific tasks, are increasingly

safetyarxiv-cs-ai
28 Apr 2026
Research

Branching Flows: Discrete, Continuous, and Manifold Flow Matching with Splits and Deletions

DGX agent

arXiv:2511.09465v3 Announce Type: replace-cross Abstract: Diffusion and flow matching approaches to generative modeling have shown promise in domains where the state space is continuous, such as image

researcharxiv-cs-lg
28 Apr 2026
Local Ai

CNN-ViT Fusion with Adaptive Attention Gate for Brain Tumor MRI Classification: A Hybrid Deep Learning Model

DGX agent

arXiv:2604.23137v1 Announce Type: cross Abstract: Early detection and classifying brain tumors using Magnetic Resonance Imaging (MRI) images is highly important but difficult to extract in medical ima

local-aiarxiv-cs-ai
28 Apr 2026
Safety

Designing Instance-Level Sampling Schedules via REINFORCE with James-Stein Shrinkage

DGX agent

arXiv:2511.22177v2 Announce Type: replace-cross Abstract: Most post-training methods for text-to-image samplers focus on model weights: either fine-tuning the backbone for alignment or distilling it f

safetyarxiv-cs-cv
28 Apr 2026
Model Releases

Diffusion Templates: A Unified Plugin Framework for Controllable Diffusion

DGX agent

arXiv:2604.24351v1 Announce Type: cross Abstract: Controllable diffusion methods have substantially expanded the practical utility of diffusion models, but they are typically developed as isolated, ba

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Don't Pause! Every prediction matters in a streaming video

DGX agent

arXiv:2604.24317v1 Announce Type: new Abstract: Streaming video models should respond the moment an event unfolds, not after the moment has passed. Yet existing online VideoQA benchmarks remain largel

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

Estimating Dense-Packed Zone Height in Liquid-Liquid Separation: A Physics-Informed Neural Network Approach

DGX agent

arXiv:2601.18399v2 Announce Type: replace Abstract: Separating liquid-liquid dispersions in gravity settlers is critical in chemical, pharmaceutical, and recycling processes. The dense-packed zone hei

model-releasesarxiv-cs-lg
28 Apr 2026
Research

Explaining Sources of Uncertainty in Automated Fact-Checking

DGX agent

arXiv:2505.17855v2 Announce Type: replace Abstract: Understanding sources of a model's uncertainty regarding its predictions is crucial for effective human-AI collaboration. Prior work proposes using

researcharxiv-cs-cl
28 Apr 2026
Model Releases

For-Value: Efficient Forward-Only Data Valuation for finetuning LLMs and VLMs

DGX agent

arXiv:2508.10180v3 Announce Type: replace Abstract: Data valuation is essential for enhancing the transparency and accountability of large language models (LLMs) and vision-language models (VLMs). How

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

Green Shielding: A User-Centric Approach Towards Trustworthy AI

DGX agent

arXiv:2604.24700v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed, yet their outputs can be highly sensitive to routine, non-adversarial variation in how users p

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

InquireMobile: Teaching VLM-based Mobile Agent to Request Human Assistance via Reinforcement Fine-Tuning

DGX agent

arXiv:2508.19679v2 Announce Type: replace Abstract: Recent advances in Vision-Language Models (VLMs) have enabled mobile agents to perceive and interact with real-world mobile environments based on hu

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

K-MetBench: A Multi-Dimensional Benchmark for Fine-Grained Evaluation of Expert Reasoning, Locality, and Multimodality in Meteorology

DGX agent

arXiv:2604.24645v1 Announce Type: cross Abstract: The development of practical (multimodal) large language model assistants for Korean weather forecasters is hindered by the absence of a multidimensio

model-releasesarxiv-cs-ai
28 Apr 2026
Research

LaDiR: Latent Diffusion Enhances LLMs for Text Reasoning

DGX agent

Large Language Models (LLMs) demonstrate their reasoning ability through chain-of-thought (CoT) generation. However, LLM’s autoregressive decoding may limit the ability to revisit and refine earlier t

researchapple-ml-research
28 Apr 2026
Applications

Latent-Hysteresis Graph ODEs: Modeling Coupled Topology-Feature Evolution via Continuous Phase Transitions

DGX agent

arXiv:2604.24293v1 Announce Type: cross Abstract: Graph neural ordinary differential equations (Graph ODEs) extend graph learning from discrete message-passing layers to continuous-time representation

applicationsarxiv-cs-ai
28 Apr 2026
Local Ai

Learning from Noisy Prompts: Saliency-Guided Prompt Distillation for Robust Segmentation with SAM

DGX agent

arXiv:2604.23314v1 Announce Type: new Abstract: Segmentation is central to clinical diagnosis and monitoring, yet the reliability of modern foundation models in medical imaging still depends on the av

local-aiarxiv-cs-cv
28 Apr 2026
Safety

Model-Free Inference of Investor Preferences: A Relative Entropy IRL Approach

DGX agent

arXiv:2604.24280v1 Announce Type: new Abstract: We present a framework using Relative Entropy Inverse Reinforcement Learning (RE-IRL) to recover investor reward functions from observed investment acti

safetyarxiv-cs-lg
28 Apr 2026
Safety

ProEval: Proactive Failure Discovery and Efficient Performance Estimation for Generative AI Evaluation

DGX agent

arXiv:2604.23099v1 Announce Type: cross Abstract: Evaluating generative AI models is increasingly resource-intensive due to slow inference, expensive raters, and a rapidly growing landscape of models

safetyarxiv-cs-ai
28 Apr 2026
Model Releases

Read the blog: https://www.together.ai/blog/together-ai-brings-nvidia-nemotron-3-nano-omni-to-developers-on-day-0#

DGX agent

Together AI announced the availability of NVIDIA Nemotron-3 Nano Omni models to developers on day zero of release, enabling early access to these multimodal AI models through their platform. The annou

model-releasestogether-ai--x
28 Apr 2026
Model Releases

Reinforcement Learning with Backtracking Feedback

DGX agent

arXiv:2602.08377v2 Announce Type: replace-cross Abstract: Addressing the critical need for robust safety in Large Language Models (LLMs), particularly against adversarial attacks and in-distribution e

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

RouteNLP: Closed-Loop LLM Routing with Conformal Cascading and Distillation Co-Optimization

DGX agent

arXiv:2604.23577v1 Announce Type: new Abstract: Serving diverse NLP workloads with large language models is costly: at one enterprise partner, inference costs exceeded $200K/month despite over 70% of

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

See Further, Think Deeper: Advancing VLM's Reasoning Ability with Low-level Visual Cues and Reflection

DGX agent

arXiv:2604.24339v1 Announce Type: cross Abstract: Recent advances in Vision-Language Models (VLMs) have benefited from Reinforcement Learning (RL) for enhanced reasoning. However, existing methods sti

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

SycoPhantasy: Quantifying Sycophancy and Hallucination in Small Open Weight VLMs for Vision-Language Scoring of Fantasy Characters

DGX agent

arXiv:2604.24346v1 Announce Type: cross Abstract: Vision-language models (VLMs) are increasingly deployed as evaluators in tasks requiring nuanced image understanding, yet their reliability in scoring

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

TextGround4M: A Prompt-Aligned Dataset for Layout-Aware Text Rendering

DGX agent

arXiv:2604.24459v1 Announce Type: new Abstract: Despite recent advances in text-to-image generation, models still struggle to accurately render prompt-specified text with correct spatial layout -- esp

model-releasesarxiv-cs-cv
28 Apr 2026
Research

Uncertainty Propagation in LLM-Based Systems

DGX agent

arXiv:2604.23505v1 Announce Type: cross Abstract: Uncertainty in large language model (LLM)-based systems is often studied at the level of a single model output, yet deployed LLM applications are comp

researcharxiv-cs-ai
28 Apr 2026
Safety

Vision-Language-Action Safety: Threats, Challenges, Evaluations, and Mechanisms

DGX agent

arXiv:2604.23775v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models are emerging as a unified substrate for embodied intelligence. This shift raises a new class of safety challenges, s

safetyarxiv-cs-ro
28 Apr 2026
Model Releases

Introducing our new work: “Learning to Orchestrate Agents in Natural Language with the Conductor” accepted at #ICLR2026 https://arxiv.org/ab…

DGX agent

Introducing our new work: “Learning to Orchestrate Agents in Natural Language with the Conductor” accepted at #ICLR2026 https://arxiv.org/abs/2512.04388 What if we trained an AI not to solve problems

model-releasesdavid-ha--x
27 Apr 2026
Research

Multi-Token Prediction via Self-Distillation

DGX agent

arXiv:2602.06019v2 Announce Type: replace Abstract: Existing techniques for accelerating language model inference, such as speculative decoding, require training auxiliary speculator models and buildi

researcharxiv-cs-cl
27 Apr 2026
Safety

PermaFrost-Attack: Stealth Pretraining Seeding(SPS) for planting Logic Landmines During LLM Training

DGX agent

arXiv:2604.22117v1 Announce Type: cross Abstract: Aligned large language models(LLMs) remain vulnerable to adversarial manipulation, and their dependence on web-scale pretraining creates a subtle but

safetyarxiv-cs-ai
27 Apr 2026
Research

Privacy Leakage via Output Label Space and Differentially Private Continual Learning

DGX agent

arXiv:2411.04680v5 Announce Type: replace Abstract: Differential privacy (DP) is a formal privacy framework that enables training machine learning (ML) models while protecting individuals' data. As po

researcharxiv-cs-lg
27 Apr 2026
Model Releases

SHAPE: Unifying Safety, Helpfulness and Pedagogy for Educational LLMs

DGX agent

arXiv:2604.22134v1 Announce Type: new Abstract: Large Language Models (LLMs) have been widely explored in educational scenarios. We identify a critical vulnerability in current educational LLMs, pedag

model-releasesarxiv-cs-cl
27 Apr 2026
Model Releases

TS-Arena -- A Live Forecast Pre-Registration Platform

DGX agent

arXiv:2512.20761v3 Announce Type: replace-cross Abstract: Time Series Foundation Models (TSFMs) are transforming the field of forecasting. However, evaluating them on historical data is increasingly d

model-releasesarxiv-cs-ai
27 Apr 2026
Research

Using Embedding Models to Improve Probabilistic Race Prediction

DGX agent

arXiv:2604.22555v1 Announce Type: new Abstract: Estimating racial disparity requires individual-level race data, which are often unavailable due to the sensitivity of collecting such information. To a

researcharxiv-cs-cl
27 Apr 2026
Model Releases

GPT-5.5 prompting guide

DGX agent

GPT-5.5 prompting guide Now that GPT-5.5 is available in the API, OpenAI have released a wealth of useful tips on how best to prompt the new model. Here's a neat trick they recommend for applications

model-releasessimon-willison
25 Apr 2026
Industry

pi needs built-in STT and TTS. @ClementDelangue what are the best open weights models that can handle all the colorful variety of non-englis…

DGX agent

This post discusses the need for built-in speech-to-text (STT) and text-to-speech (TTS) capabilities in a product or platform (likely referring to Hugging Face's platform based on the mention of Clem

industryclem-delangue--x
25 Apr 2026
Model Releases

ADS-POI: Agentic Spatiotemporal State Decomposition for Next Point-of-Interest Recommendation

DGX agent

arXiv:2604.20846v1 Announce Type: cross Abstract: Next point-of-interest (POI) recommendation requires modeling user mobility as a spatiotemporal sequence, where different behavioral factors may evolv

model-releasesarxiv-cs-ai
24 Apr 2026
Research

Analytical FFN-to-MoE Restructuring via Activation Pattern Analysis

DGX agent

arXiv:2502.04416v3 Announce Type: replace-cross Abstract: Scaling large language models (LLMs) improves performance but significantly increases inference costs, with feed-forward networks (FFNs) consu

researcharxiv-cs-ai
24 Apr 2026
Model Releases

Breaking Bad: Interpretability-Based Safety Audits of State-of-the-Art LLMs

DGX agent

arXiv:2604.20945v1 Announce Type: cross Abstract: Effective safety auditing of large language models (LLMs) demands tools that go beyond black-box probing and systematically uncover vulnerabilities ro

model-releasesarxiv-cs-lg
24 Apr 2026
Research

Clinically-Informed Modeling for Pediatric Brain Tumor Classification from Whole-Slide Histopathology Images

DGX agent

arXiv:2604.21060v1 Announce Type: new Abstract: Accurate diagnosis of pediatric brain tumors, starting with histopathology, presents unique challenges for deep learning, including severe data scarcity

researcharxiv-cs-cv
24 Apr 2026
Model Releases

GeoRA: Geometry-Aware Low-Rank Adaptation for RLVR

DGX agent

arXiv:2601.09361v3 Announce Type: replace-cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) is a key paradigm for improving large-scale reasoning models. Unlike supervised fine-tun

model-releasesarxiv-cs-ai
24 Apr 2026
Safety

How VLAs (Really) Work In Open-World Environments

DGX agent

arXiv:2604.21192v1 Announce Type: cross Abstract: Vision-language-action models (VLAs) have been extensively used in robotics applications, achieving great success in various manipulation problems. Mo

safetyarxiv-cs-ai
24 Apr 2026
Model Releases

Hyperloop Transformers

DGX agent

arXiv:2604.21254v1 Announce Type: cross Abstract: LLM architecture research generally aims to maximize model quality subject to fixed compute/latency budgets. However, many applications of interest su

model-releasesarxiv-cs-cl
24 Apr 2026
← Previous
1…343344345346347…1337
Next →