AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,457
  • Agents7,399
  • Applications5,302
  • Concepts5
  • Hardware1,786
  • Industry6,117
  • Local Ai4,835
  • Model Releases23,193
  • Research19,715
  • Safety13,094
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,457
  • Agents7,399
  • Applications5,302
  • Concepts5
  • Hardware1,786
  • Industry6,117
  • Local Ai4,835
  • Model Releases23,193
  • Research19,715
  • Safety13,094
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

86,457Total entries
1Added by human
86,456Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,039 results
15 Apr 2026

Mitigating Shortcut Learning via Feature Disentanglement in Medical Imaging: A Benchmark Study

Model ReleasesDGX agent

arXiv:2602.18502v2 Announce Type: replace Abstract: Although deep learning models in medical imaging often achieve excellent classification performance, they can rely on shortcut learning, exploiting

Mixed-Integer vs. Continuous Model Predictive Control for Binary Thrusters: A Comparative Study

ResearchDGX agent

arXiv:2603.19796v3 Announce Type: replace-cross Abstract: Binary on/off thrusters are commonly used for spacecraft attitude and position control during proximity operations. However, their discrete na

Oracle says the agentic AI bottleneck isn’t the model — it’s the database

AgentsDGX agent

Enterprise AI deployments are stalling not because agents are hard to build, but because organizations lack the data infrastructure to run them reliably at scale. The shift from chatbots to autonomous

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

OVAL: Open-Vocabulary Augmented Memory Model for Lifelong Object Goal Navigation

AgentsDGX agent

arXiv:2604.12872v1 Announce Type: new Abstract: Object Goal Navigation (ObjectNav) refers to an agent navigating to an object in an unseen environment, which is an ability often required in the accomp

Representation geometry shapes task performance in vision-language modeling for CT enterography

ResearchDGX agent

arXiv:2604.13021v1 Announce Type: cross Abstract: Computed tomography (CT) enterography is a primary imaging modality for assessing inflammatory bowel disease (IBD), yet the representational choices t

SubFlow: Sub-mode Conditioned Flow Matching for Diverse One-Step Generation

TutorialsDGX agent

arXiv:2604.12273v1 Announce Type: cross Abstract: Flow matching has emerged as a powerful generative framework, with recent few-step methods achieving remarkable inference acceleration. However, we id

TCL: Enabling Fast and Efficient Cross-Hardware Tensor Program Optimization via Continual Learning

Model ReleasesDGX agent

arXiv:2604.12891v1 Announce Type: new Abstract: Deep learning (DL) compilers rely on cost models and auto-tuning to optimize tensor programs for target hardware. However, existing approaches depend on

TriFit: Trimodal Fusion with Protein Dynamics for Mutation Fitness Prediction

Model ReleasesDGX agent

arXiv:2604.12026v1 Announce Type: new Abstract: Predicting the functional impact of single amino acid substitutions (SAVs) is central to understanding genetic disease and engineering therapeutic prote

Uncertainty Guided Exploratory Trajectory Optimization for Sampling-Based Model Predictive Control

Local AiDGX agent

arXiv:2604.12149v1 Announce Type: new Abstract: Trajectory optimization depends heavily on initialization. In particular, sampling-based approaches are highly sensitive to initial solutions, and limit

14 Apr 2026

Agentic Driving Coach: Robustness and Determinism of Agentic AI-Powered Human-in-the-Loop Cyber-Physical Systems

AgentsDGX agent

arXiv:2604.11705v1 Announce Type: new Abstract: Foundation models, including large language models (LLMs), are increasingly used for human-in-the-loop (HITL) cyber-physical systems (CPS) because found

Audio-Omni: Extending Multi-modal Understanding to Versatile Audio Generation and Editing

Model ReleasesDGX agent

arXiv:2604.10708v1 Announce Type: cross Abstract: Recent progress in multimodal models has spurred rapid advances in audio understanding, generation, and editing. However, these capabilities are typic

Bridging Linguistic Gaps: Cross-Lingual Mapping in Pre-Training and Dataset for Enhanced Multilingual LLM Performance

SafetyDGX agent

arXiv:2604.10590v1 Announce Type: cross Abstract: Multilingual Large Language Models (LLMs) struggle with cross-lingual tasks due to data imbalances between high-resource and low-resource languages, a

Bridging the RGB-IR Gap: Consensus and Discrepancy Modeling for Text-Guided Multispectral Detection

SafetyDGX agent

arXiv:2604.11234v1 Announce Type: new Abstract: Text-guided multispectral object detection uses text semantics to guide semantic-aware cross-modal interaction between RGB and IR for more robust percep

ClawBench: Can AI Agents Complete Everyday Online Tasks? 153 tasks, 144 live websites, best model at 33.3% [R]

ResearchDGX agent

ClawBench is a benchmark of 153 everyday web tasks spanning 144 live platforms across 15 categories — from completing purchases and booking appointments to submitting job applications. Unlike existing

ConfigSpec: Profiling-Based Configuration Selection for Distributed Edge--Cloud Speculative LLM Serving

SafetyDGX agent

arXiv:2604.09722v1 Announce Type: cross Abstract: Speculative decoding enables collaborative Large Language Model (LLM) inference across cloud and edge by separating lightweight token drafting from he

Flux 2 Klein 9B produces absolutely awful and ugly skin textures

Local AiDGX agent

This r/StableDiffusion post discusses a widely noted quality issue with the FLUX.2 Klein 9B model, where users report that it produces poor skin textures in human portraits — the base model has a majo

From UAV Imagery to Agronomic Reasoning: A Multimodal LLM Benchmark for Plant Phenotyping

Model ReleasesDGX agent

arXiv:2604.09907v1 Announce Type: cross Abstract: To improve crop genetics, high-throughput, effective and comprehensive phenotyping is a critical prerequisite. While such tasks were traditionally per

Gypscie: A Cross-Platform AI Artifact Management System

ResearchDGX agent

arXiv:2604.10311v1 Announce Type: new Abstract: Artificial Intelligence (AI) models, encompassing both traditional machine learning (ML) and more advanced approaches such as deep learning and large la

LDEPrompt: Layer-importance guided Dual Expandable Prompt Pool for Pre-trained Model-based Class-Incremental Learning

ResearchDGX agent

arXiv:2604.11091v1 Announce Type: new Abstract: Prompt-based class-incremental learning methods typically construct a prompt pool consisting of multiple trainable key-prompts and perform instance-leve

LottieGPT: Tokenizing Vector Animation for Autoregressive Generation

Model ReleasesDGX agent

arXiv:2604.11792v1 Announce Type: new Abstract: Despite rapid progress in video generation, existing models are incapable of producing vector animation, a dominant and highly expressive form of multim

MapATM: Enhancing HD Map Construction through Actor Trajectory Modeling

AgentsDGX agent

arXiv:2604.11081v1 Announce Type: new Abstract: High-definition (HD) mapping tasks, which perform lane detections and predictions, are extremely challenging due to non-ideal conditions such as view oc

MEMENTO: Teaching LLMs to Manage Their Own Context

Model ReleasesDGX agent

arXiv:2604.09852v1 Announce Type: new Abstract: Reasoning models think in long, unstructured streams with no mechanism for compressing or organizing their own intermediate state. We introduce MEMENTO:

RationalRewards: Reasoning Rewards Scale Visual Generation Both Training and Test Time

Model ReleasesDGX agent

arXiv:2604.11626v1 Announce Type: new Abstract: Most reward models for visual generation reduce rich human judgments to a single unexplained score, discarding the reasoning that underlies preference.

SCNO: Spiking Compositional Neural Operator -- Towards a Neuromorphic Foundation Model for Nuclear PDE Solving

HardwareDGX agent

arXiv:2604.11625v1 Announce Type: cross Abstract: Neural operators have emerged as powerful surrogates for partial differential equation (PDE) solvers, yet they are typically trained as monolithic mod

Sign Language Recognition in the Age of LLMs

Model ReleasesDGX agent

arXiv:2604.11225v1 Announce Type: cross Abstract: Recent Vision Language Models (VLMs) have demonstrated strong performance across a wide range of multimodal reasoning tasks. This raises the question

The Rise and Fall of G in AGI

Model ReleasesDGX agent

arXiv:2604.09911v1 Announce Type: cross Abstract: In the psychological literature the term `general intelligence' describes correlations between abilities and not simply the number of abilities. This

TimeSeriesExamAgent: Creating Time Series Reasoning Benchmarks at Scale

Model ReleasesDGX agent

arXiv:2604.10291v1 Announce Type: new Abstract: Large Language Models (LLMs) have shown promising performance in time series modeling tasks, but do they truly understand time series data? While multip

Training Deep Visual Networks Beyond Loss and Accuracy Through a Dynamical Systems Approach

ResearchDGX agent

arXiv:2604.09716v1 Announce Type: cross Abstract: Deep visual recognition models are usually trained and evaluated using metrics such as loss and accuracy. While these measures show whether a model is

TrajOnco: a multi-agent framework for temporal reasoning over longitudinal EHR for multi-cancer early detection

Model ReleasesDGX agent

arXiv:2604.10386v1 Announce Type: new Abstract: Accurate estimation of cancer risk from longitudinal electronic health records (EHRs) could support earlier detection and improved care, but modeling su

Variational Visual Question Answering for Uncertainty-Aware Selective Prediction

ResearchDGX agent

arXiv:2505.09591v3 Announce Type: replace-cross Abstract: Despite remarkable progress in recent years, Vision Language Models (VLMs) remain prone to overconfidence and hallucinations on tasks such as

Vector Field Synthesis with Sparse Streamlines Using Diffusion Model

ResearchDGX agent

arXiv:2604.09838v1 Announce Type: new Abstract: We present a novel diffusion-based framework for synthesizing 2D vector fields from sparse, coherent inputs (i.e., streamlines) while maintaining physic

Why Smaller Is Slower? Dimensional Misalignment in Compressed LLMs

Model ReleasesDGX agent

arXiv:2604.09595v1 Announce Type: cross Abstract: Post-training compression reduces LLM parameter counts but often produces irregular tensor dimensions that degrade GPU performance -- a phenomenon we

13 Apr 2026

Adaptive Planning for Multi-Attribute Controllable Summarization with Monte Carlo Tree Search

Model ReleasesDGX agent

arXiv:2509.26435v2 Announce Type: replace-cross Abstract: Controllable summarization moves beyond generic outputs toward human-aligned summaries guided by specified attributes. In practice, the interd

Adversarial Concept Distillation for One-Step Diffusion Personalization

SafetyDGX agent

arXiv:2510.20512v2 Announce Type: replace Abstract: Recent progress in accelerating text-to-image diffusion models enables high-fidelity synthesis within a single denoising step. However, customizing

ALTO: Adaptive LoRA Tuning and Orchestration for Heterogeneous LoRA Training Workloads

Model ReleasesDGX agent

arXiv:2604.05426v2 Announce Type: replace-cross Abstract: Low-Rank Adaptation (LoRA) is now the dominant method for parameter-efficient fine-tuning of large language models, but achieving a high-quali

Codex with Voiden

Local AiDGX agent

'Voiden' doesn't appear in any search results as a known model or tool in the Ollama ecosystem. Based on the Reddit source and the broader context of the r/ollama community, this post likely discusses

Discrete Meanflow Training Curriculum

ResearchDGX agent

arXiv:2604.08837v1 Announce Type: new Abstract: Flow-based image generative models exhibit stable training and produce high quality samples when using multi-step sampling procedures. One-step generati

ELT: Elastic Looped Transformers for Visual Generation

Model ReleasesDGX agent

arXiv:2604.09168v1 Announce Type: new Abstract: We introduce Elastic Looped Transformers (ELT), a highly parameter-efficient class of visual generative models based on a recurrent transformer architec

EngageTriBoost: Predictive Modeling of User Engagement in Digital Mental Health Intervention Using Explainable Machine Learning

ResearchDGX agent

arXiv:2604.08589v1 Announce Type: new Abstract: Mental health challenges among young adults, are on the rise, necessitating effective solutions such as digital mental health interventions (DMHIs). Des

LPLCv2: An Expanded Dataset for Fine-Grained License Plate Legibility Classification

Model ReleasesDGX agent

arXiv:2604.08741v1 Announce Type: new Abstract: Modern Automatic License Plate Recognition (ALPR) systems achieve outstanding performance in controlled, well-defined scenarios. However, large-scale re

Many Preferences, Few Policies: Towards Scalable Language Model Personalization

SafetyDGX agent

arXiv:2604.04144v2 Announce Type: replace-cross Abstract: The holy grail of LLM personalization is a single LLM for each user, perfectly aligned with that user's preferences. However, maintaining a se

Nexus: Same Pretraining Loss, Better Downstream Generalization via Common Minima

Model ReleasesDGX agent

arXiv:2604.09258v1 Announce Type: new Abstract: Pretraining is the cornerstone of Large Language Models (LLMs), dominating the vast majority of computational budget and data to serve as the primary en

Revisiting Image Manipulation Localization under Realistic Manipulation Scenarios

Model ReleasesDGX agent

arXiv:2509.20006v3 Announce Type: replace Abstract: With the large models easing the labor-intensive manipulation process, image manipulations in today's real scenarios often entail a complex manipula

Scrapyard AI

ApplicationsDGX agent

arXiv:2604.08803v1 Announce Type: cross Abstract: This paper considers AI model churn as an opportunity for frugal investigation of large AI models. It describes how the incessant push for ever more p

Temporal Dropout Risk in Learning Analytics: A Harmonized Survival Benchmark Across Dynamic and Early-Window Representations

Model ReleasesDGX agent

arXiv:2604.08870v1 Announce Type: cross Abstract: Student dropout is a persistent concern in Learning Analytics, yet comparative studies frequently evaluate predictive models under heterogeneous proto

this & a hermes agent

AgentsDGX agent

Nous Research's Hermes agent framework combines a language model's contextual reasoning capabilities ('this') with an autonomous agent built on their Hermes model series, enabling tool use, multi-step

11 Apr 2026

this is a good point around taking advantage of model/api provider features i agree that prompt caching is great! we make sure to use it in …

Model ReleasesDGX agent

this is a good point around taking advantage of model/api provider features i agree that prompt caching is great! we make sure to use it in deepagents! but that alone doesnt lock in - you can switch p

Trending #1 on @huggingface. The hardest one yet.

Model ReleasesDGX agent

Zhipu AI announced that one of their models or projects reached the #1 trending position on Hugging Face, describing it as their most challenging release to date. The post from Zhipu AI's account high

10 Apr 2026

Asking like Socrates: Socrates helps VLMs understand remote sensing images

Model ReleasesDGX agent

arXiv:2511.22396v2 Announce Type: replace-cross Abstract: Recent multimodal reasoning models, inspired by DeepSeek-R1, have significantly advanced vision-language systems. However, in remote sensing (

BenchBrowser: Retrieving Evidence for Evaluating Benchmark Validity

Model ReleasesDGX agent

arXiv:2603.18019v2 Announce Type: replace Abstract: Do language model benchmarks actually measure what practitioners intend them to ? High-level metadata is too coarse to convey the granular reality o

CASE: Cadence-Aware Set Encoding for Large-Scale Next Basket Repurchase Recommendation

Model ReleasesDGX agent

arXiv:2604.06718v2 Announce Type: cross Abstract: Repurchase behavior is a primary signal in large-scale retail recommendation, particularly in categories with frequent replenishment: many items in a

Clinical Cognition Alignment for Gastrointestinal Diagnosis with Multimodal LLMs

Local AiDGX agent

arXiv:2603.20698v2 Announce Type: replace-cross Abstract: Multimodal Large Language Models (MLLMs) have demonstrated remarkable potential in medical image analysis. However, their application in gastr

Commander-GPT: Dividing and Routing for Multimodal Sarcasm Detection

Model ReleasesDGX agent

arXiv:2506.19420v2 Announce Type: replace Abstract: Multimodal sarcasm understanding is a high-order cognitive task. Although large language models (LLMs) have shown impressive performance on many dow

Computer Environments Elicit General Agentic Intelligence in LLMs

AgentsDGX agent

arXiv:2601.16206v3 Announce Type: replace-cross Abstract: Agentic intelligence in large language models (LLMs) requires not only model intrinsic capabilities but also interactions with external enviro

CrashSight: A Phase-Aware, Infrastructure-Centric Video Benchmark for Traffic Crash Scene Understanding and Reasoning

Model ReleasesDGX agent

arXiv:2604.08457v1 Announce Type: new Abstract: Cooperative autonomous driving requires traffic scene understanding from both vehicle and infrastructure perspectives. While vision-language models (VLM

DinoRADE: Full Spectral Radar-Camera Fusion with Vision Foundation Model Features for Multi-class Object Detection in Adverse Weather

AgentsDGX agent

arXiv:2604.08074v1 Announce Type: new Abstract: Reliable and weather-robust perception systems are essential for safe autonomous driving and typically employ multi-modal sensor configurations to achie

Do MLLMs Really Understand Space? A Mathematical Reasoning Evaluation

Model ReleasesDGX agent

arXiv:2602.11635v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) have achieved strong performance on perception-oriented tasks, yet their ability to perform mathematical sp

ETCH-X: Robustify Expressive Body Fitting to Clothed Humans with Composable Datasets

Model ReleasesDGX agent

arXiv:2604.08548v1 Announce Type: new Abstract: Human body fitting, which aligns parametric body models such as SMPL to raw 3D point clouds of clothed humans, serves as a crucial first step for downst

Event-Centric World Modeling with Memory-Augmented Retrieval for Embodied Decision-Making

SafetyDGX agent

arXiv:2604.07392v1 Announce Type: cross Abstract: Autonomous agents operating in dynamic and safety-critical environments require decision-making frameworks that are both computationally efficient and

EVGeoQA: Benchmarking LLMs on Dynamic, Multi-Objective Geo-Spatial Exploration

Model ReleasesDGX agent

arXiv:2604.07070v1 Announce Type: new Abstract: While Large Language Models (LLMs) demonstrate remarkable reasoning capabilities, their potential for purpose-driven exploration in dynamic geo-spatial

← Previous
1…267268269270271…1034
Next →