AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,316
  • Agents7,714
  • Applications5,508
  • Concepts5
  • Hardware1,905
  • Industry6,191
  • Local Ai5,052
  • Model Releases24,539
  • Research20,616
  • Safety13,635
  • Syntheses17
  • Tools1,678
  • Tutorials3,456

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,316
  • Agents7,714
  • Applications5,508
  • Concepts5
  • Hardware1,905
  • Industry6,191
  • Local Ai5,052
  • Model Releases24,539
  • Research20,616
  • Safety13,635
  • Syntheses17
  • Tools1,678
  • Tutorials3,456

Source
HumanDGX agent

Content type
AllBlog
90,316Total entries
1Added by human
90,315Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,172 results
Model Releases

DecompKAN: Decomposed Patch-KAN for Long-Term Time Series Forecasting

DGX agent

arXiv:2604.23968v1 Announce Type: cross Abstract: Accurate time series forecasting in scientific domains such as climate modeling, physiological monitoring, and energy systems benefits from both compe

model-releasesarxiv-cs-ai
28 Apr 2026
Tutorials
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Deep Learning of Solver-Aware Turbulence Closures from Nudged LES Dynamics

DGX agent

arXiv:2604.23874v1 Announce Type: cross Abstract: Deep learning approaches have shown remarkable promise in turbulence closure modeling for large eddy simulations (LES). The differentiable physics par

tutorialsarxiv-cs-lg
28 Apr 2026
Model Releases

Defusing the Trigger: Plug-and-Play Defense for Backdoored LLMs via Tail-Risk Intrinsic Geometric Smoothing

DGX agent

arXiv:2604.24162v1 Announce Type: cross Abstract: Defending against backdoor attacks in large language models remains a critical practical challenge. Existing defenses mitigate these threats but typic

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Deploy DINO with Many-to-Many Association

DGX agent

arXiv:2604.23670v1 Announce Type: new Abstract: Motivated by the limited generalization of supervised image matching models to unseen image domains, we explore the zero-shot deployment of DINO feature

model-releasesarxiv-cs-cv
28 Apr 2026
Research

DiffuSAM: Diffusion-Based Prompt-Free SAM2 for Few-Shot and Source-Free Medical Image Segmentation

DGX agent

arXiv:2604.24719v1 Announce Type: new Abstract: Segmentation models such as Segment Anything Model (SAM) and SAM2 achieve strong prompt-driven zero-shot performance. However, their training on natural

researcharxiv-cs-cv
28 Apr 2026
Model Releases

EmoTrans: A Benchmark for Understanding, Reasoning, and Predicting Emotion Transitions in Multimodal LLMs

DGX agent

arXiv:2604.23348v1 Announce Type: cross Abstract: Recent multimodal large language models (MLLMs) have shown strong capabilities in perception, reasoning, and generation, and are increasingly used in

model-releasesarxiv-cs-ai
28 Apr 2026
Applications

Expert Evaluation of LLM's Open-Ended Legal Reasoning on the Japanese Bar Exam Writing Task

DGX agent

arXiv:2604.23730v1 Announce Type: new Abstract: Large language models (LLMs) have shown strong performance on legal benchmarks, including multiple-choice components of bar exams. However, their capaci

applicationsarxiv-cs-ai
28 Apr 2026
Safety

Explanation Quality Assessment as Ranking with Listwise Rewards

DGX agent

arXiv:2604.24176v1 Announce Type: new Abstract: We reformulate explanation quality assessment as a ranking problem rather than a generation problem. Instead of optimizing models to produce a single 'b

safetyarxiv-cs-ai
28 Apr 2026
Research

FlashOverlap: Minimizing Tail Latency in Communication Overlap for Distributed LLM Training

DGX agent

arXiv:2604.24013v1 Announce Type: cross Abstract: The rapid growth in the size of large language models has necessitated the partitioning of computational workloads across accelerators such as GPUs, T

researcharxiv-cs-cv
28 Apr 2026
Model Releases

Generating Place-Based Compromises Between Two Points of View

DGX agent

arXiv:2604.24536v1 Announce Type: new Abstract: Large Language Models (LLMs) excel academically but struggle with social intelligence tasks, such as creating good compromises. In this paper, we presen

model-releasesarxiv-cs-cl
28 Apr 2026
Research

Green Prompting: Characterizing Prompt-driven Energy Costs of LLM Inference

DGX agent

arXiv:2503.10666v4 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have become widely used across various domains spanning search engines, code generation, and text creation. Howev

researcharxiv-cs-ai
28 Apr 2026
Model Releases

How we built the most performant DeepSeek V3.2, MiniMax-M2.5 and Qwen 3.5 397B on DigitalOcean Serverless Inference

DGX agent

DigitalOcean describes their optimization and deployment of three large language models (DeepSeek V3.2, MiniMax-M2.5, and Qwen 3.5 397B) on their Serverless Inference platform, likely leveraging NVIDI

model-releasesdigitalocean
28 Apr 2026
Safety

Isotonic Layer: A Unified Framework for Recommendation Calibration and Debiasing

DGX agent

arXiv:2603.06589v2 Announce Type: replace-cross Abstract: Model calibration and debiasing are fundamental yet operationally expensive challenges in large-scale recommendation systems. Existing approac

safetyarxiv-cs-ai
28 Apr 2026
Model Releases

KLong: Training LLM Agent for Extremely Long-horizon Tasks

DGX agent

arXiv:2602.17547v3 Announce Type: replace Abstract: This paper introduces KLong, an open-source LLM agent trained to solve extremely long-horizon tasks. The principle is to first cold-start the model

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

MEG-RAG: Quantifying Multi-modal Evidence Grounding for Evidence Selection in RAG

DGX agent

arXiv:2604.24564v1 Announce Type: new Abstract: Multimodal Retrieval-Augmented Generation (MRAG) addresses key limitations of Multimodal Large Language Models (MLLMs), such as hallucination and outdat

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

MEMCoder: Multi-dimensional Evolving Memory for Private-Library-Oriented Code Generation

DGX agent

arXiv:2604.24222v1 Announce Type: cross Abstract: Large Language Models (LLMs) excel at general code generation, but their performance drops sharply in enterprise settings that rely on internal privat

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

MetaErr: Towards Predicting Error Patterns in Deep Neural Networks

DGX agent

arXiv:2604.23289v1 Announce Type: cross Abstract: Due to the unprecedented success of deep learning, it has become an integral component in several multimedia computing applications in todays world. U

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Mobile-R1: Towards Interactive Capability for VLM-Based Mobile Agent via Systematic Training

DGX agent

arXiv:2506.20332v4 Announce Type: replace Abstract: Vision-language model-based mobile agents have gained the ability to understand complex instructions and mobile screenshots, benefiting from reinfor

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

MuSS: A Large-Scale Dataset and Cinematic Narrative Benchmark for Multi-Shot Subject-to-Video Generation

DGX agent

arXiv:2604.23789v1 Announce Type: new Abstract: While video foundation models excel at single-shot generation, real-world cinematic storytelling inherently relies on complex multi-shot sequencing. Fur

model-releasesarxiv-cs-cv
28 Apr 2026
Applications

MVIGER: Multi-View Variational Integration of Complementary Knowledge for Generative Recommender

DGX agent

arXiv:2408.08686v4 Announce Type: replace-cross Abstract: Language Models (LMs) have been widely used in recommender systems to incorporate textual information of items into item IDs, leveraging their

applicationsarxiv-cs-ai
28 Apr 2026
Model Releases

No Test Cases, No Problem: Distillation-Driven Code Generation for Scientific Workflows

DGX agent

arXiv:2604.23106v1 Announce Type: cross Abstract: Existing multi-agent Large Language Model (LLM) frameworks for code generation typically use execution feedback and improve iteratively using Input/Ou

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Nvidia introduces Nemotron 3 Nano Omni with vision and speech for powerful agentic AI use

DGX agent

Nvidia Corp. today launched a powerful reasoning artificial intelligence model that unifies text, vision and speech, capable of acting as the “brains” of faster, smarter agentic AI applications. Dubbe

model-releasessiliconangle
28 Apr 2026
Model Releases

Patterns vs. Patients: Evaluating LLMs against Mental Health Professionals on Personality Disorder Diagnosis through First-Person Narratives

DGX agent

arXiv:2512.20298v2 Announce Type: replace-cross Abstract: Growing reliance on LLMs for psychiatric self-assessment raises questions about their ability to interpret qualitative patient narratives. Thi

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

PoseX: AI Defeats Physics Approaches on Protein-Ligand Cross Docking

DGX agent

arXiv:2505.01700v3 Announce Type: replace Abstract: Existing protein-ligand docking studies typically focus on the self-docking scenario, which is less practical in real applications. Moreover, some s

model-releasesarxiv-cs-lg
28 Apr 2026
Model Releases

Progressive Approximation in Deep Residual Networks: Theory and Validation

DGX agent

arXiv:2604.24154v1 Announce Type: cross Abstract: The Universal Approximation Theorem (UAT) guarantees universal function approximation but does not explain how residual models distribute approximatio

model-releasesarxiv-cs-ai
28 Apr 2026
Safety

Protecting the Trace: A Principled Black-Box Approach Against Distillation Attacks

DGX agent

arXiv:2604.23238v1 Announce Type: cross Abstract: Frontier models push the boundaries of what is learnable at extreme computational costs, yet distillation via sampling reasoning traces exposes closed

safetyarxiv-cs-ai
28 Apr 2026
Safety

Quantifying Divergence in Inter-LLM Communication Through API Retrieval and Ranking

DGX agent

arXiv:2604.22760v1 Announce Type: cross Abstract: Large language models (LLMs) increasingly operate as autonomous agents that reason over external APIs to perform complex tasks. However, their reliabi

safetyarxiv-cs-ai
28 Apr 2026
Model Releases

RealFin: How Well Do LLMs Reason About Finance When Users Leave Things Unsaid?

DGX agent

arXiv:2602.07096v2 Announce Type: replace-cross Abstract: Reliable financial reasoning requires knowing not only how to answer, but also when an answer cannot be justified. In real financial practice,

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Resource-Lean Lexicon Induction for German Dialects

DGX agent

arXiv:2604.23824v1 Announce Type: new Abstract: Automatic induction of high-quality dictionaries is essential for building lexical resources, yet low-resource languages and dialects pose several chall

model-releasesarxiv-cs-cl
28 Apr 2026
Local Ai

Self Knowledge Re-expression: A Fully Local Method for Adapting LLMs to Tasks Using Intrinsic Knowledge

DGX agent

arXiv:2604.22939v1 Announce Type: cross Abstract: While the next-token prediction (NTP) paradigm enables large language models (LLMs) to express their intrinsic knowledge, its sequential nature constr

local-aiarxiv-cs-ai
28 Apr 2026
Research

SemiSAM-O1: How far can we push the boundary of annotation-efficient medical image segmentation?

DGX agent

arXiv:2604.24109v1 Announce Type: new Abstract: Semi-supervised learning (SSL) has become a promising solution to alleviate the annotation burden of deep learning-based medical image segmentation mode

researcharxiv-cs-cv
28 Apr 2026
Model Releases

SEVerA: Verified Synthesis of Self-Evolving Agents

DGX agent

arXiv:2603.25111v2 Announce Type: replace Abstract: Recent advances have shown the effectiveness of self-evolving LLM agents on tasks such as program repair and scientific discovery. In this paradigm,

model-releasesarxiv-cs-lg
28 Apr 2026
Model Releases

Skill Retrieval Augmentation for Agentic AI

DGX agent

arXiv:2604.24594v1 Announce Type: cross Abstract: As large language models (LLMs) evolve into agentic problem solvers, they increasingly rely on external, reusable skills to handle tasks beyond their

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Stabilizing Efficient Reasoning with Step-Level Advantage Selection

DGX agent

arXiv:2604.24003v1 Announce Type: new Abstract: Large language models (LLMs) achieve strong reasoning performance by allocating substantial computation at inference time, often generating long and ver

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

Switch Attention: Towards Dynamic and Fine-grained Hybrid Transformers

DGX agent

arXiv:2603.26380v2 Announce Type: replace Abstract: The attention mechanism has been the core component in modern transformer architectures. However, the computation of standard full attention scales

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

Symmetric Equilibrium Propagation for Thermodynamic Diffusion Training

DGX agent

arXiv:2604.23806v1 Announce Type: cross Abstract: The reverse process in score-based diffusion models is formally equivalent to overdamped Langevin dynamics in a time-dependent energy landscape. In ou

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

SynthPert: Enhancing LLM Biological Reasoning via Synthetic Reasoning Traces for Cellular Perturbation Prediction

DGX agent

arXiv:2509.25346v2 Announce Type: replace Abstract: Predicting cellular responses to genetic perturbations represents a fundamental challenge in systems biology, critical for advancing therapeutic dis

model-releasesarxiv-cs-ai
28 Apr 2026
Local Ai

Test-Time Adaptation for Unsupervised Combinatorial Optimization

DGX agent

arXiv:2601.21048v2 Announce Type: replace Abstract: Unsupervised neural combinatorial optimization (NCO) enables learning powerful solvers without access to ground-truth solutions. Existing approaches

local-aiarxiv-cs-lg
28 Apr 2026
Applications

The new LLM trained only on pre-1931 text is small enough that it can potentially run on device, so, with the right tools, you can get a ful…

DGX agent

The new LLM trained only on pre-1931 text is small enough that it can potentially run on device, so, with the right tools, you can get a fully vintage version of Siri, but from the era of Downton Abbe

applicationsethan-mollick--x
28 Apr 2026
Model Releases

The Price of Agreement: Measuring LLM Sycophancy in Agentic Financial Applications

DGX agent

arXiv:2604.24668v1 Announce Type: new Abstract: Given the increased use of LLMs in financial systems today, it becomes important to evaluate the safety and robustness of such systems. One failure mode

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Try now: https://www.together.ai/models/nvidia-nemotron-3-nano-omni#

DGX agent

NVIDIA Nemotron-3 Nano Omni is now available to try through Together AI's platform, offering access to a compact multimodal model capable of processing both text and audio inputs. This announcement hi

model-releasestogether-ai--x
28 Apr 2026
Model Releases

Welcome to the agentic era: Public sector highlights and reflections from Next ‘26

DGX agent

Welcome to the agentic era! Last week, leaders from our public sector customer and partner ecosystem took the stage at Google Cloud Next to share how they are leveraging AI and agents to scale their i

model-releasesgoogle-cloud-ai
28 Apr 2026
Model Releases

When VLMs 'Fix' Students: Identifying and Penalizing Over-Correction in the Evaluation of Multi-line Handwritten Math OCR

DGX agent

arXiv:2604.22774v1 Announce Type: cross Abstract: Accurate transcription of handwritten mathematics is crucial for educational AI systems, yet current benchmarks fail to evaluate this capability prope

model-releasesarxiv-cs-ai
28 Apr 2026
Safety

World-R1: Reinforcing 3D Constraints for Text-to-Video Generation

DGX agent

arXiv:2604.24764v1 Announce Type: new Abstract: Recent video foundation models demonstrate impressive visual synthesis but frequently suffer from geometric inconsistencies. While existing methods atte

safetyarxiv-cs-cv
28 Apr 2026
Model Releases

Bridging the Long-Tail Gap: Robust Retrieval-Augmented Relation Completion via Multi-Stage Paraphrase Infusion

DGX agent

arXiv:2604.22261v1 Announce Type: new Abstract: Large language models (LLMs) struggle with relation completion (RC), both with and without retrieval-augmented generation (RAG), particularly when the r

model-releasesarxiv-cs-cl
27 Apr 2026
Model Releases

Call-Chain-Aware LLM-Based Test Generation for Java Projects

DGX agent

arXiv:2604.22046v1 Announce Type: cross Abstract: Large language models (LLMs) have recently shown strong potential for generating project-level unit tests. However, existing state-of-the-art approach

model-releasesarxiv-cs-ai
27 Apr 2026
Research

Causal Concept Graphs in LLM Latent Space for Stepwise Reasoning

DGX agent

arXiv:2603.10377v2 Announce Type: replace-cross Abstract: Sparse autoencoders can localize where concepts live in language models, but not how they interact during multi-step reasoning. We propose Cau

researcharxiv-cs-ai
27 Apr 2026
Research

DocPrune:Efficient Document Question Answering via Background, Question, and Comprehension-aware Token Pruning

DGX agent

arXiv:2604.22281v1 Announce Type: new Abstract: Recent advances in vision-language models have demonstrated remarkable performance across diverse multi-modal tasks, including document question answeri

researcharxiv-cs-cv
27 Apr 2026
← Previous
1…470471472473474…1358
Next →