AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,428
  • Agents7,398
  • Applications5,301
  • Concepts5
  • Hardware1,785
  • Industry6,113
  • Local Ai4,833
  • Model Releases23,177
  • Research19,713
  • Safety13,092
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,428
  • Agents7,398
  • Applications5,301
  • Concepts5
  • Hardware1,785
  • Industry6,113
  • Local Ai4,833
  • Model Releases23,177
  • Research19,713
  • Safety13,092
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

86,428Total entries
1Added by human
86,427Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,016 results
1 May 2026

AutoREC: A software platform for developing reinforcement learning agents for equivalent circuit model generation from electrochemical impedance spectroscopy data

AgentsDGX agent

arXiv:2604.27266v1 Announce Type: new Abstract: This paper introduces AutoREC, an open-source Python package for developing reinforcement learning (RL) agents to automatically generate equivalent circ

AutoVDC: Automated Vision Data Cleaning Using Vision-Language Models

AgentsDGX agent

arXiv:2507.12414v2 Announce Type: replace-cross Abstract: Training of autonomous driving systems requires extensive datasets with precise annotations to attain robust performance. Human annotations su

FluxMoE: Decoupling Expert Residency for High-Performance MoE Serving

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.02715v2 Announce Type: replace Abstract: Mixture-of-Experts (MoE) models have become a dominant paradigm for scaling large language models, but their rapidly growing parameter sizes introdu

Nebius acquires AI model optimization startup Eigen AI for $643M

ApplicationsDGX agent

Nebius Group NV, a Dutch operator of artificial intelligence data centers, today announced plans to buy software maker Eigen AI Inc. for 643 million. The company will finance the acquisition with cash

PhyCo: Learning Controllable Physical Priors for Generative Motion

Model ReleasesDGX agent

arXiv:2604.28169v1 Announce Type: cross Abstract: Modern video diffusion models excel at appearance synthesis but still struggle with physical consistency: objects drift, collisions lack realistic reb

Rethinking Agentic Reinforcement Learning In Large Language Models

AgentsDGX agent

arXiv:2604.27859v1 Announce Type: new Abstract: Reinforcement Learning (RL) has traditionally focused on training specialized agents to optimize predefined reward functions within narrowly defined env

Simulating Validity: Modal Decoupling in MLLM Generated Feedback on Science Drawings

Model ReleasesDGX agent

arXiv:2604.26957v1 Announce Type: cross Abstract: In science education, students frequently construct hand-drawn visual models of scientific phenomena. These drawings rely on a visual structure where

TopBench: A Benchmark for Implicit Prediction and Reasoning over Tabular Question Answering

Model ReleasesDGX agent

arXiv:2604.28076v1 Announce Type: cross Abstract: Large Language Models (LLMs) have advanced Table Question Answering, where most queries can be answered by extracting information or simple aggregatio

30 Apr 2026

Cloud CISO Perspectives: At Next ‘26, why we’re multicloud and multi-AI

Model ReleasesDGX agent

Welcome to the second Cloud CISO Perspectives for April 2026. Today, Francis deSouza, COO Google Cloud and President, Security Products, explains why Google is multicloud and multi-AI, straight from N

Comparative Analysis of AutoML and BiLSTM Models for Cyberbullying Detection on Indonesian Instagram Comments

ResearchDGX agent

arXiv:2604.26229v1 Announce Type: new Abstract: This study compares machine learning and deep learning approaches for cyberbullying detection in Indonesian-language Instagram comments. Using a balance

Electricity price forecasting across Norway's five bidding zones in the post-crisis era

Model ReleasesDGX agent

arXiv:2604.26634v1 Announce Type: new Abstract: Norway's electricity market is heavily dominated by hydropower, but the 2021--2022 energy crisis and stronger integration with Continental Europe have f

EvolvingAgent: Curriculum Self-evolving Agent with Continual World Model for Long-Horizon Tasks

AgentsDGX agent

arXiv:2502.05907v3 Announce Type: replace Abstract: Completing Long-Horizon (LH) tasks in open-ended worlds is an important yet difficult problem for embodied agents. Existing approaches suffer from t

Motion-Driven Multi-Object Tracking of Model Organisms in Space Science Experiments

ResearchDGX agent

arXiv:2604.26321v1 Announce Type: new Abstract: Automated animal behavior analysis relies on long-term, interpretable individual trajectories; however, multi-animal tracking in space science experimen

OmniTrend: Content-Context Modeling for Scalable Social Popularity Prediction

ResearchDGX agent

arXiv:2604.26252v1 Announce Type: new Abstract: Predicting social media popularity requires understanding both the intrinsic appeal of content and the external context that determines how it is expose

Talent or Luck? Evaluating Attribution Bias in Large Language Models

SafetyDGX agent

arXiv:2505.22910v2 Announce Type: replace Abstract: When a student fails an exam, do we tend to blame their effort or the test's difficulty? Attribution, defined as how reasons are assigned to event o

Tell Model Where to Look: Mitigating Hallucinations in MLLMs by Vision-Guided Attention

Local AiDGX agent

arXiv:2511.20032v3 Announce Type: replace Abstract: Visual attention serves as the primary mechanism through which MLLMs interpret visual information; however, its limited localization capability ofte

What Kind of Language is Easy to Language-Model Under Curriculum Learning?

SafetyDGX agent

arXiv:2604.26844v1 Announce Type: new Abstract: Many of the thousands of attested languages share common configurations of features, creating a spectrum from typologically very rare (e.g., object-verb

29 Apr 2026

From Syntax to Emotion: A Mechanistic Analysis of Emotion Inference in LLMs

ResearchDGX agent

arXiv:2604.25866v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used in emotionally sensitive human-AI applications, yet little is known about how emotion recognition is

MGSM-Pro: A Simple Strategy for Robust Multilingual Mathematical Reasoning Evaluation

Model ReleasesDGX agent

arXiv:2601.21225v2 Announce Type: replace Abstract: Large language models have made substantial progress in mathematical reasoning. However, benchmark development for multilingual evaluation has lagge

Mutual Forcing: Dual-Mode Self-Evolution for Fast Autoregressive Audio-Video Character Generation

ResearchDGX agent

arXiv:2604.25819v1 Announce Type: new Abstract: In this work, we propose Mutual Forcing, a framework for fast autoregressive audio-video generation with long-horizon audio-video synchronization. Our a

Novel 3D Binary Indexed Tree for Volume Computation of 3D Reconstructed Models from Volumetric Data

ResearchDGX agent

arXiv:2412.10441v2 Announce Type: replace-cross Abstract: In the burgeoning field of medical imaging, precise computation of 3D volume holds a significant importance for subsequent qualitative analysi

OMHBench: Benchmarking Balanced and Grounded Omni-Modal Multi-Hop Reasoning

Model ReleasesDGX agent

arXiv:2508.16198v3 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) have increasingly supported omni-modal processing across text, vision, and speech. However, existing evalua

One Perturbation, Two Failure Modes: Probing VLM Safety via Embedding-Guided Typographic Perturbations

Model ReleasesDGX agent

arXiv:2604.25102v1 Announce Type: new Abstract: Typographic prompt injection exploits vision language models' (VLMs) ability to read text rendered in images, posing a growing threat as VLMs power auto

Quantum-Inspired Robust and Scalable SAR Object Classification

ResearchDGX agent

arXiv:2604.25755v1 Announce Type: cross Abstract: SAR image classification naturally has to deal with huge noise and a high dynamic range particularly requiring robust classification models. Additiona

28 Apr 2026

A Self-Supervised Framework for Space Object Behaviour Characterisation

SafetyDGX agent

arXiv:2504.06176v3 Announce Type: replace-cross Abstract: Foundation Models, which leverage large neural networks pre-trained on unlabelled data before fine-tuning for specific tasks, are increasingly

Branching Flows: Discrete, Continuous, and Manifold Flow Matching with Splits and Deletions

ResearchDGX agent

arXiv:2511.09465v3 Announce Type: replace-cross Abstract: Diffusion and flow matching approaches to generative modeling have shown promise in domains where the state space is continuous, such as image

CNN-ViT Fusion with Adaptive Attention Gate for Brain Tumor MRI Classification: A Hybrid Deep Learning Model

Local AiDGX agent

arXiv:2604.23137v1 Announce Type: cross Abstract: Early detection and classifying brain tumors using Magnetic Resonance Imaging (MRI) images is highly important but difficult to extract in medical ima

Designing Instance-Level Sampling Schedules via REINFORCE with James-Stein Shrinkage

SafetyDGX agent

arXiv:2511.22177v2 Announce Type: replace-cross Abstract: Most post-training methods for text-to-image samplers focus on model weights: either fine-tuning the backbone for alignment or distilling it f

Diffusion Templates: A Unified Plugin Framework for Controllable Diffusion

Model ReleasesDGX agent

arXiv:2604.24351v1 Announce Type: cross Abstract: Controllable diffusion methods have substantially expanded the practical utility of diffusion models, but they are typically developed as isolated, ba

Don't Pause! Every prediction matters in a streaming video

Model ReleasesDGX agent

arXiv:2604.24317v1 Announce Type: new Abstract: Streaming video models should respond the moment an event unfolds, not after the moment has passed. Yet existing online VideoQA benchmarks remain largel

Estimating Dense-Packed Zone Height in Liquid-Liquid Separation: A Physics-Informed Neural Network Approach

Model ReleasesDGX agent

arXiv:2601.18399v2 Announce Type: replace Abstract: Separating liquid-liquid dispersions in gravity settlers is critical in chemical, pharmaceutical, and recycling processes. The dense-packed zone hei

Explaining Sources of Uncertainty in Automated Fact-Checking

ResearchDGX agent

arXiv:2505.17855v2 Announce Type: replace Abstract: Understanding sources of a model's uncertainty regarding its predictions is crucial for effective human-AI collaboration. Prior work proposes using

For-Value: Efficient Forward-Only Data Valuation for finetuning LLMs and VLMs

Model ReleasesDGX agent

arXiv:2508.10180v3 Announce Type: replace Abstract: Data valuation is essential for enhancing the transparency and accountability of large language models (LLMs) and vision-language models (VLMs). How

Green Shielding: A User-Centric Approach Towards Trustworthy AI

Model ReleasesDGX agent

arXiv:2604.24700v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed, yet their outputs can be highly sensitive to routine, non-adversarial variation in how users p

InquireMobile: Teaching VLM-based Mobile Agent to Request Human Assistance via Reinforcement Fine-Tuning

Model ReleasesDGX agent

arXiv:2508.19679v2 Announce Type: replace Abstract: Recent advances in Vision-Language Models (VLMs) have enabled mobile agents to perceive and interact with real-world mobile environments based on hu

K-MetBench: A Multi-Dimensional Benchmark for Fine-Grained Evaluation of Expert Reasoning, Locality, and Multimodality in Meteorology

Model ReleasesDGX agent

arXiv:2604.24645v1 Announce Type: cross Abstract: The development of practical (multimodal) large language model assistants for Korean weather forecasters is hindered by the absence of a multidimensio

LaDiR: Latent Diffusion Enhances LLMs for Text Reasoning

ResearchDGX agent

Large Language Models (LLMs) demonstrate their reasoning ability through chain-of-thought (CoT) generation. However, LLM’s autoregressive decoding may limit the ability to revisit and refine earlier t

Latent-Hysteresis Graph ODEs: Modeling Coupled Topology-Feature Evolution via Continuous Phase Transitions

ApplicationsDGX agent

arXiv:2604.24293v1 Announce Type: cross Abstract: Graph neural ordinary differential equations (Graph ODEs) extend graph learning from discrete message-passing layers to continuous-time representation

Learning from Noisy Prompts: Saliency-Guided Prompt Distillation for Robust Segmentation with SAM

Local AiDGX agent

arXiv:2604.23314v1 Announce Type: new Abstract: Segmentation is central to clinical diagnosis and monitoring, yet the reliability of modern foundation models in medical imaging still depends on the av

Model-Free Inference of Investor Preferences: A Relative Entropy IRL Approach

SafetyDGX agent

arXiv:2604.24280v1 Announce Type: new Abstract: We present a framework using Relative Entropy Inverse Reinforcement Learning (RE-IRL) to recover investor reward functions from observed investment acti

ProEval: Proactive Failure Discovery and Efficient Performance Estimation for Generative AI Evaluation

SafetyDGX agent

arXiv:2604.23099v1 Announce Type: cross Abstract: Evaluating generative AI models is increasingly resource-intensive due to slow inference, expensive raters, and a rapidly growing landscape of models

Read the blog: https://www.together.ai/blog/together-ai-brings-nvidia-nemotron-3-nano-omni-to-developers-on-day-0#

Model ReleasesDGX agent

Together AI announced the availability of NVIDIA Nemotron-3 Nano Omni models to developers on day zero of release, enabling early access to these multimodal AI models through their platform. The annou

Reinforcement Learning with Backtracking Feedback

Model ReleasesDGX agent

arXiv:2602.08377v2 Announce Type: replace-cross Abstract: Addressing the critical need for robust safety in Large Language Models (LLMs), particularly against adversarial attacks and in-distribution e

RouteNLP: Closed-Loop LLM Routing with Conformal Cascading and Distillation Co-Optimization

Model ReleasesDGX agent

arXiv:2604.23577v1 Announce Type: new Abstract: Serving diverse NLP workloads with large language models is costly: at one enterprise partner, inference costs exceeded $200K/month despite over 70% of

See Further, Think Deeper: Advancing VLM's Reasoning Ability with Low-level Visual Cues and Reflection

Model ReleasesDGX agent

arXiv:2604.24339v1 Announce Type: cross Abstract: Recent advances in Vision-Language Models (VLMs) have benefited from Reinforcement Learning (RL) for enhanced reasoning. However, existing methods sti

SycoPhantasy: Quantifying Sycophancy and Hallucination in Small Open Weight VLMs for Vision-Language Scoring of Fantasy Characters

Model ReleasesDGX agent

arXiv:2604.24346v1 Announce Type: cross Abstract: Vision-language models (VLMs) are increasingly deployed as evaluators in tasks requiring nuanced image understanding, yet their reliability in scoring

TextGround4M: A Prompt-Aligned Dataset for Layout-Aware Text Rendering

Model ReleasesDGX agent

arXiv:2604.24459v1 Announce Type: new Abstract: Despite recent advances in text-to-image generation, models still struggle to accurately render prompt-specified text with correct spatial layout -- esp

Uncertainty Propagation in LLM-Based Systems

ResearchDGX agent

arXiv:2604.23505v1 Announce Type: cross Abstract: Uncertainty in large language model (LLM)-based systems is often studied at the level of a single model output, yet deployed LLM applications are comp

Vision-Language-Action Safety: Threats, Challenges, Evaluations, and Mechanisms

SafetyDGX agent

arXiv:2604.23775v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models are emerging as a unified substrate for embodied intelligence. This shift raises a new class of safety challenges, s

27 Apr 2026

Introducing our new work: “Learning to Orchestrate Agents in Natural Language with the Conductor” accepted at #ICLR2026 https://arxiv.org/ab…

Model ReleasesDGX agent

Introducing our new work: “Learning to Orchestrate Agents in Natural Language with the Conductor” accepted at #ICLR2026 https://arxiv.org/abs/2512.04388 What if we trained an AI not to solve problems

Multi-Token Prediction via Self-Distillation

ResearchDGX agent

arXiv:2602.06019v2 Announce Type: replace Abstract: Existing techniques for accelerating language model inference, such as speculative decoding, require training auxiliary speculator models and buildi

PermaFrost-Attack: Stealth Pretraining Seeding(SPS) for planting Logic Landmines During LLM Training

SafetyDGX agent

arXiv:2604.22117v1 Announce Type: cross Abstract: Aligned large language models(LLMs) remain vulnerable to adversarial manipulation, and their dependence on web-scale pretraining creates a subtle but

Privacy Leakage via Output Label Space and Differentially Private Continual Learning

ResearchDGX agent

arXiv:2411.04680v5 Announce Type: replace Abstract: Differential privacy (DP) is a formal privacy framework that enables training machine learning (ML) models while protecting individuals' data. As po

SHAPE: Unifying Safety, Helpfulness and Pedagogy for Educational LLMs

Model ReleasesDGX agent

arXiv:2604.22134v1 Announce Type: new Abstract: Large Language Models (LLMs) have been widely explored in educational scenarios. We identify a critical vulnerability in current educational LLMs, pedag

TS-Arena -- A Live Forecast Pre-Registration Platform

Model ReleasesDGX agent

arXiv:2512.20761v3 Announce Type: replace-cross Abstract: Time Series Foundation Models (TSFMs) are transforming the field of forecasting. However, evaluating them on historical data is increasingly d

Using Embedding Models to Improve Probabilistic Race Prediction

ResearchDGX agent

arXiv:2604.22555v1 Announce Type: new Abstract: Estimating racial disparity requires individual-level race data, which are often unavailable due to the sensitivity of collecting such information. To a

25 Apr 2026

GPT-5.5 prompting guide

Model ReleasesDGX agent

GPT-5.5 prompting guide Now that GPT-5.5 is available in the API, OpenAI have released a wealth of useful tips on how best to prompt the new model. Here's a neat trick they recommend for applications

pi needs built-in STT and TTS. @ClementDelangue what are the best open weights models that can handle all the colorful variety of non-englis…

IndustryDGX agent

This post discusses the need for built-in speech-to-text (STT) and text-to-speech (TTS) capabilities in a product or platform (likely referring to Hugging Face's platform based on the mention of Clem

24 Apr 2026

ADS-POI: Agentic Spatiotemporal State Decomposition for Next Point-of-Interest Recommendation

Model ReleasesDGX agent

arXiv:2604.20846v1 Announce Type: cross Abstract: Next point-of-interest (POI) recommendation requires modeling user mobility as a spatiotemporal sequence, where different behavioral factors may evolv

Analytical FFN-to-MoE Restructuring via Activation Pattern Analysis

ResearchDGX agent

arXiv:2502.04416v3 Announce Type: replace-cross Abstract: Scaling large language models (LLMs) improves performance but significantly increases inference costs, with feed-forward networks (FFNs) consu

← Previous
1…264265266267268…1034
Next →