AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries89,023
  • Agents7,608
  • Applications5,444
  • Concepts5
  • Hardware1,854
  • Industry6,179
  • Local Ai4,968
  • Model Releases24,142
  • Research20,258
  • Safety13,456
  • Syntheses17
  • Tools1,677
  • Tutorials3,415

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries89,023
  • Agents7,608
  • Applications5,444
  • Concepts5
  • Hardware1,854
  • Industry6,179
  • Local Ai4,968
  • Model Releases24,142
  • Research20,258
  • Safety13,456
  • Syntheses17
  • Tools1,677
  • Tutorials3,415

Source
HumanDGX agent

Content type
89,023Total entries
1Added by human
89,022Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
64,154 results
Model Releases

LongCoT: Benchmarking Long-Horizon Chain-of-Thought Reasoning

DGX agent

arXiv:2604.14140v1 Announce Type: new Abstract: As language models are increasingly deployed for complex autonomous tasks, their ability to reason accurately over longer horizons becomes critical. An

model-releasesarxiv-cs-lg
16 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

MedRCube: A Multidimensional Framework for Fine-Grained and In-Depth Evaluation of MLLMs in Medical Imaging

DGX agent

arXiv:2604.13756v1 Announce Type: new Abstract: The potential of Multimodal Large Language Models (MLLMs) in domain of medical imaging raise the demands of systematic and rigorous evaluation framework

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Multi-Task LLM with LoRA Fine-Tuning for Automated Cancer Staging and Biomarker Extraction

DGX agent

arXiv:2604.13328v1 Announce Type: new Abstract: Pathology reports serve as the definitive record for breast cancer staging, yet their unstructured format impedes large-scale data curation. While Large

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

Synthesizing Instruction-Tuning Datasets with Contrastive Decoding

DGX agent

arXiv:2604.13538v1 Announce Type: new Abstract: Using responses generated by high-performing large language models (LLMs) for instruction tuning has become a widely adopted approach. However, the exis

model-releasesarxiv-cs-cl
16 Apr 2026
Research

Training-Free Semantic Multi-Object Tracking with Vision-Language Models

DGX agent

arXiv:2604.14074v1 Announce Type: new Abstract: Semantic Multi-Object Tracking (SMOT) extends multi-object tracking with semantic outputs such as video summaries, instance-level captions, and interact

researcharxiv-cs-cv
16 Apr 2026
Model Releases

TREX: Automating LLM Fine-tuning via Agent-Driven Tree-based Exploration

DGX agent

arXiv:2604.14116v1 Announce Type: cross Abstract: While Large Language Models (LLMs) have empowered AI research agents to perform isolated scientific tasks, automating complex, real-world workflows, s

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

ViBES: A Conversational Agent with Behaviorally-Intelligent 3D Virtual Body

DGX agent

arXiv:2512.14234v2 Announce Type: replace Abstract: Human communication is inherently multimodal and social: words, prosody, and body language jointly carry intent. Yet most prior systems model human

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

Beyond Scores: Diagnostic LLM Evaluation via Fine-Grained Abilities

DGX agent

arXiv:2604.12191v1 Announce Type: new Abstract: Current evaluations of large language models aggregate performance across diverse tasks into single scores. This obscures fine-grained ability variation

model-releasesarxiv-cs-ai
15 Apr 2026
Safety

Causal Diffusion Models for Counterfactual Outcome Distributions in Longitudinal Data

DGX agent

arXiv:2604.12992v1 Announce Type: cross Abstract: Predicting counterfactual outcomes in longitudinal data, where sequential treatment decisions heavily depend on evolving patient states, is critical y

safetyarxiv-cs-lg
15 Apr 2026
Model Releases

Comparison of low Steps, Klein 9b x Z image turbo x Ernie Turbo x Qwen 2512 8 Steps

DGX agent

This r/StableDiffusion post presents a community-driven visual comparison of several modern, fast text-to-image diffusion models — Flux.2 Klein 9B (a smaller, faster distillation of Flux.2 Dev availab

model-releasesr-stablediffusion
15 Apr 2026
Local Ai

CREG: Compass Relational Evidence Graph for Characterizing Directional Structure in VLM Spatial-Reasoning Attribution

DGX agent

arXiv:2603.20475v3 Announce Type: replace Abstract: Standard attribution heatmaps show where a vision-language model (VLM) focuses, but they do not reveal whether the recovered evidence is organized b

local-aiarxiv-cs-cv
15 Apr 2026
Research

Does Visual Token Pruning Improve Calibration? An Empirical Study on Confidence in MLLMs

DGX agent

arXiv:2604.12035v1 Announce Type: new Abstract: Visual token pruning is a widely used strategy for efficient inference in multimodal large language models (MLLMs), but existing work mainly evaluates i

researcharxiv-cs-cv
15 Apr 2026
Model Releases

Dual-Modality Anchor-Guided Filtering for Test-time Prompt Tuning

DGX agent

arXiv:2604.12403v1 Announce Type: new Abstract: Test-Time Prompt Tuning (TPT) adapts vision-language models using augmented views, but its effectiveness is hindered by the challenge of determining whi

model-releasesarxiv-cs-cv
15 Apr 2026
Model Releases

Efficient Adversarial Training via Criticality-Aware Fine-Tuning

DGX agent

arXiv:2604.12780v1 Announce Type: cross Abstract: Vision Transformer (ViT) models have achieved remarkable performance across various vision tasks, with scalability being a key advantage when applied

model-releasesarxiv-cs-ai
15 Apr 2026
Applications

GGD-SLAM: Monocular 3DGS SLAM Powered by Generalizable Motion Model for Dynamic Environments

DGX agent

arXiv:2604.12837v1 Announce Type: new Abstract: Visual SLAM algorithms achieve significant improvements through the exploration of 3D Gaussian Splatting (3DGS) representations, particularly in generat

applicationsarxiv-cs-ro
15 Apr 2026
Applications

Interpretable Relational Inference with LLM-Guided Symbolic Dynamics Modeling

DGX agent

arXiv:2604.12806v1 Announce Type: new Abstract: Inferring latent interaction structures from observed dynamics is a fundamental inverse problem in many-body interacting systems. Most neural approaches

applicationsarxiv-cs-lg
15 Apr 2026
Model Releases

Is Gemma 4 26B MoE or 31B good as an MCP agent for coding with Xcode?

DGX agent

This r/ollama discussion explores the suitability of Google's Gemma 4 models — specifically the 26B Mixture of Experts (MoE) and 31B Dense variants — as MCP (Model Context Protocol) agents for coding

model-releasesr-ollama
15 Apr 2026
Safety

Meet Dynamic Individual Preferences: Resolving Conflicting Human Value with Paired Fine-Tuning

DGX agent

arXiv:2604.12479v1 Announce Type: new Abstract: Recent advances in large language models (LLMs) have significantly improved the alignment of models with general human preferences. However, a major cha

safetyarxiv-cs-cl
15 Apr 2026
Model Releases

Mitigating Shortcut Learning via Feature Disentanglement in Medical Imaging: A Benchmark Study

DGX agent

arXiv:2602.18502v2 Announce Type: replace Abstract: Although deep learning models in medical imaging often achieve excellent classification performance, they can rely on shortcut learning, exploiting

model-releasesarxiv-cs-cv
15 Apr 2026
Research

Mixed-Integer vs. Continuous Model Predictive Control for Binary Thrusters: A Comparative Study

DGX agent

arXiv:2603.19796v3 Announce Type: replace-cross Abstract: Binary on/off thrusters are commonly used for spacecraft attitude and position control during proximity operations. However, their discrete na

researcharxiv-cs-ro
15 Apr 2026
Agents

Oracle says the agentic AI bottleneck isn’t the model — it’s the database

DGX agent

Enterprise AI deployments are stalling not because agents are hard to build, but because organizations lack the data infrastructure to run them reliably at scale. The shift from chatbots to autonomous

agentssiliconangle
15 Apr 2026
Agents

OVAL: Open-Vocabulary Augmented Memory Model for Lifelong Object Goal Navigation

DGX agent

arXiv:2604.12872v1 Announce Type: new Abstract: Object Goal Navigation (ObjectNav) refers to an agent navigating to an object in an unseen environment, which is an ability often required in the accomp

agentsarxiv-cs-ro
15 Apr 2026
Research

Representation geometry shapes task performance in vision-language modeling for CT enterography

DGX agent

arXiv:2604.13021v1 Announce Type: cross Abstract: Computed tomography (CT) enterography is a primary imaging modality for assessing inflammatory bowel disease (IBD), yet the representational choices t

researcharxiv-cs-ai
15 Apr 2026
Tutorials

SubFlow: Sub-mode Conditioned Flow Matching for Diverse One-Step Generation

DGX agent

arXiv:2604.12273v1 Announce Type: cross Abstract: Flow matching has emerged as a powerful generative framework, with recent few-step methods achieving remarkable inference acceleration. However, we id

tutorialsarxiv-cs-cv
15 Apr 2026
Model Releases

TCL: Enabling Fast and Efficient Cross-Hardware Tensor Program Optimization via Continual Learning

DGX agent

arXiv:2604.12891v1 Announce Type: new Abstract: Deep learning (DL) compilers rely on cost models and auto-tuning to optimize tensor programs for target hardware. However, existing approaches depend on

model-releasesarxiv-cs-lg
15 Apr 2026
Model Releases

TriFit: Trimodal Fusion with Protein Dynamics for Mutation Fitness Prediction

DGX agent

arXiv:2604.12026v1 Announce Type: new Abstract: Predicting the functional impact of single amino acid substitutions (SAVs) is central to understanding genetic disease and engineering therapeutic prote

model-releasesarxiv-cs-lg
15 Apr 2026
Local Ai

Uncertainty Guided Exploratory Trajectory Optimization for Sampling-Based Model Predictive Control

DGX agent

arXiv:2604.12149v1 Announce Type: new Abstract: Trajectory optimization depends heavily on initialization. In particular, sampling-based approaches are highly sensitive to initial solutions, and limit

local-aiarxiv-cs-ro
15 Apr 2026
Agents

Agentic Driving Coach: Robustness and Determinism of Agentic AI-Powered Human-in-the-Loop Cyber-Physical Systems

DGX agent

arXiv:2604.11705v1 Announce Type: new Abstract: Foundation models, including large language models (LLMs), are increasingly used for human-in-the-loop (HITL) cyber-physical systems (CPS) because found

agentsarxiv-cs-ai
14 Apr 2026
Model Releases

Audio-Omni: Extending Multi-modal Understanding to Versatile Audio Generation and Editing

DGX agent

arXiv:2604.10708v1 Announce Type: cross Abstract: Recent progress in multimodal models has spurred rapid advances in audio understanding, generation, and editing. However, these capabilities are typic

model-releasesarxiv-cs-ai
14 Apr 2026
Safety

Bridging Linguistic Gaps: Cross-Lingual Mapping in Pre-Training and Dataset for Enhanced Multilingual LLM Performance

DGX agent

arXiv:2604.10590v1 Announce Type: cross Abstract: Multilingual Large Language Models (LLMs) struggle with cross-lingual tasks due to data imbalances between high-resource and low-resource languages, a

safetyarxiv-cs-ai
14 Apr 2026
Safety

Bridging the RGB-IR Gap: Consensus and Discrepancy Modeling for Text-Guided Multispectral Detection

DGX agent

arXiv:2604.11234v1 Announce Type: new Abstract: Text-guided multispectral object detection uses text semantics to guide semantic-aware cross-modal interaction between RGB and IR for more robust percep

safetyarxiv-cs-cv
14 Apr 2026
Research

ClawBench: Can AI Agents Complete Everyday Online Tasks? 153 tasks, 144 live websites, best model at 33.3% [R]

DGX agent

ClawBench is a benchmark of 153 everyday web tasks spanning 144 live platforms across 15 categories — from completing purchases and booking appointments to submitting job applications. Unlike existing

researchr-machinelearning
14 Apr 2026
Safety

ConfigSpec: Profiling-Based Configuration Selection for Distributed Edge--Cloud Speculative LLM Serving

DGX agent

arXiv:2604.09722v1 Announce Type: cross Abstract: Speculative decoding enables collaborative Large Language Model (LLM) inference across cloud and edge by separating lightweight token drafting from he

safetyarxiv-cs-ai
14 Apr 2026
Local Ai

Flux 2 Klein 9B produces absolutely awful and ugly skin textures

DGX agent

This r/StableDiffusion post discusses a widely noted quality issue with the FLUX.2 Klein 9B model, where users report that it produces poor skin textures in human portraits — the base model has a majo

local-air-stablediffusion
14 Apr 2026
Model Releases

From UAV Imagery to Agronomic Reasoning: A Multimodal LLM Benchmark for Plant Phenotyping

DGX agent

arXiv:2604.09907v1 Announce Type: cross Abstract: To improve crop genetics, high-throughput, effective and comprehensive phenotyping is a critical prerequisite. While such tasks were traditionally per

model-releasesarxiv-cs-ai
14 Apr 2026
Research

Gypscie: A Cross-Platform AI Artifact Management System

DGX agent

arXiv:2604.10311v1 Announce Type: new Abstract: Artificial Intelligence (AI) models, encompassing both traditional machine learning (ML) and more advanced approaches such as deep learning and large la

researcharxiv-cs-ai
14 Apr 2026
Research

LDEPrompt: Layer-importance guided Dual Expandable Prompt Pool for Pre-trained Model-based Class-Incremental Learning

DGX agent

arXiv:2604.11091v1 Announce Type: new Abstract: Prompt-based class-incremental learning methods typically construct a prompt pool consisting of multiple trainable key-prompts and perform instance-leve

researcharxiv-cs-cv
14 Apr 2026
Model Releases

LottieGPT: Tokenizing Vector Animation for Autoregressive Generation

DGX agent

arXiv:2604.11792v1 Announce Type: new Abstract: Despite rapid progress in video generation, existing models are incapable of producing vector animation, a dominant and highly expressive form of multim

model-releasesarxiv-cs-cv
14 Apr 2026
Agents

MapATM: Enhancing HD Map Construction through Actor Trajectory Modeling

DGX agent

arXiv:2604.11081v1 Announce Type: new Abstract: High-definition (HD) mapping tasks, which perform lane detections and predictions, are extremely challenging due to non-ideal conditions such as view oc

agentsarxiv-cs-cv
14 Apr 2026
Model Releases

MEMENTO: Teaching LLMs to Manage Their Own Context

DGX agent

arXiv:2604.09852v1 Announce Type: new Abstract: Reasoning models think in long, unstructured streams with no mechanism for compressing or organizing their own intermediate state. We introduce MEMENTO:

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

RationalRewards: Reasoning Rewards Scale Visual Generation Both Training and Test Time

DGX agent

arXiv:2604.11626v1 Announce Type: new Abstract: Most reward models for visual generation reduce rich human judgments to a single unexplained score, discarding the reasoning that underlies preference.

model-releasesarxiv-cs-ai
14 Apr 2026
Hardware

SCNO: Spiking Compositional Neural Operator -- Towards a Neuromorphic Foundation Model for Nuclear PDE Solving

DGX agent

arXiv:2604.11625v1 Announce Type: cross Abstract: Neural operators have emerged as powerful surrogates for partial differential equation (PDE) solvers, yet they are typically trained as monolithic mod

hardwarearxiv-cs-ai
14 Apr 2026
Model Releases

Sign Language Recognition in the Age of LLMs

DGX agent

arXiv:2604.11225v1 Announce Type: cross Abstract: Recent Vision Language Models (VLMs) have demonstrated strong performance across a wide range of multimodal reasoning tasks. This raises the question

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

The Rise and Fall of G in AGI

DGX agent

arXiv:2604.09911v1 Announce Type: cross Abstract: In the psychological literature the term `general intelligence' describes correlations between abilities and not simply the number of abilities. This

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

TimeSeriesExamAgent: Creating Time Series Reasoning Benchmarks at Scale

DGX agent

arXiv:2604.10291v1 Announce Type: new Abstract: Large Language Models (LLMs) have shown promising performance in time series modeling tasks, but do they truly understand time series data? While multip

model-releasesarxiv-cs-ai
14 Apr 2026
Research

Training Deep Visual Networks Beyond Loss and Accuracy Through a Dynamical Systems Approach

DGX agent

arXiv:2604.09716v1 Announce Type: cross Abstract: Deep visual recognition models are usually trained and evaluated using metrics such as loss and accuracy. While these measures show whether a model is

researcharxiv-cs-ai
14 Apr 2026
Model Releases

TrajOnco: a multi-agent framework for temporal reasoning over longitudinal EHR for multi-cancer early detection

DGX agent

arXiv:2604.10386v1 Announce Type: new Abstract: Accurate estimation of cancer risk from longitudinal electronic health records (EHRs) could support earlier detection and improved care, but modeling su

model-releasesarxiv-cs-ai
14 Apr 2026
Research

Variational Visual Question Answering for Uncertainty-Aware Selective Prediction

DGX agent

arXiv:2505.09591v3 Announce Type: replace-cross Abstract: Despite remarkable progress in recent years, Vision Language Models (VLMs) remain prone to overconfidence and hallucinations on tasks such as

researcharxiv-cs-ai
14 Apr 2026
← Previous
1…346347348349350…1337
Next →