AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,573
  • Agents7,495
  • Applications5,364
  • Concepts5
  • Hardware1,813
  • Industry6,149
  • Local Ai4,892
  • Model Releases23,569
  • Research19,967
  • Safety13,263
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,573
  • Agents7,495
  • Applications5,364
  • Concepts5
  • Hardware1,813
  • Industry6,149
  • Local Ai4,892
  • Model Releases23,569
  • Research19,967
  • Safety13,263
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
HumanDGX agent

87,573Total entries
1Added by human
87,572Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,952 results
28 Apr 2026

Computational Design and Co-Robotic Fabrication for Material Reuse in Architecture

ApplicationsDGX agent

arXiv:2604.24648v1 Announce Type: new Abstract: Climate change and resource depletion demand a shift from the dominant linear 'take-make-use-dispose' paradigm of construction toward circular, low-wast

Continual Calibration: Coverage Can Collapse Before Accuracy in Lifelong LLM Fine-Tuning

ResearchDGX agent

arXiv:2604.23987v1 Announce Type: new Abstract: Continual learning for large language models is typically evaluated through accuracy retention under sequential fine-tuning. We argue that this perspect

Cortex-Inspired Continual Learning: Unsupervised Instantiation and Recovery of Functional Task Networks

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.24637v1 Announce Type: cross Abstract: Block-sequential continual learning demands that a single model both protect prior solutions from catastrophic forgetting and efficiently infer at inf

Coverage-Based Calibration for Post-Training Quantization via Weighted Set Cover over Outlier Channels

Model ReleasesDGX agent

arXiv:2604.24008v1 Announce Type: new Abstract: Post-Training Quantization (PTQ) compresses large language models to low bit-widths using a small calibration set, and its quality depends strongly on w

DecompKAN: Decomposed Patch-KAN for Long-Term Time Series Forecasting

Model ReleasesDGX agent

arXiv:2604.23968v1 Announce Type: cross Abstract: Accurate time series forecasting in scientific domains such as climate modeling, physiological monitoring, and energy systems benefits from both compe

Deep Learning of Solver-Aware Turbulence Closures from Nudged LES Dynamics

TutorialsDGX agent

arXiv:2604.23874v1 Announce Type: cross Abstract: Deep learning approaches have shown remarkable promise in turbulence closure modeling for large eddy simulations (LES). The differentiable physics par

Defusing the Trigger: Plug-and-Play Defense for Backdoored LLMs via Tail-Risk Intrinsic Geometric Smoothing

Model ReleasesDGX agent

arXiv:2604.24162v1 Announce Type: cross Abstract: Defending against backdoor attacks in large language models remains a critical practical challenge. Existing defenses mitigate these threats but typic

Deploy DINO with Many-to-Many Association

Model ReleasesDGX agent

arXiv:2604.23670v1 Announce Type: new Abstract: Motivated by the limited generalization of supervised image matching models to unseen image domains, we explore the zero-shot deployment of DINO feature

DiffuSAM: Diffusion-Based Prompt-Free SAM2 for Few-Shot and Source-Free Medical Image Segmentation

ResearchDGX agent

arXiv:2604.24719v1 Announce Type: new Abstract: Segmentation models such as Segment Anything Model (SAM) and SAM2 achieve strong prompt-driven zero-shot performance. However, their training on natural

EmoTrans: A Benchmark for Understanding, Reasoning, and Predicting Emotion Transitions in Multimodal LLMs

Model ReleasesDGX agent

arXiv:2604.23348v1 Announce Type: cross Abstract: Recent multimodal large language models (MLLMs) have shown strong capabilities in perception, reasoning, and generation, and are increasingly used in

Expert Evaluation of LLM's Open-Ended Legal Reasoning on the Japanese Bar Exam Writing Task

ApplicationsDGX agent

arXiv:2604.23730v1 Announce Type: new Abstract: Large language models (LLMs) have shown strong performance on legal benchmarks, including multiple-choice components of bar exams. However, their capaci

Explanation Quality Assessment as Ranking with Listwise Rewards

SafetyDGX agent

arXiv:2604.24176v1 Announce Type: new Abstract: We reformulate explanation quality assessment as a ranking problem rather than a generation problem. Instead of optimizing models to produce a single 'b

FlashOverlap: Minimizing Tail Latency in Communication Overlap for Distributed LLM Training

ResearchDGX agent

arXiv:2604.24013v1 Announce Type: cross Abstract: The rapid growth in the size of large language models has necessitated the partitioning of computational workloads across accelerators such as GPUs, T

Generating Place-Based Compromises Between Two Points of View

Model ReleasesDGX agent

arXiv:2604.24536v1 Announce Type: new Abstract: Large Language Models (LLMs) excel academically but struggle with social intelligence tasks, such as creating good compromises. In this paper, we presen

Green Prompting: Characterizing Prompt-driven Energy Costs of LLM Inference

ResearchDGX agent

arXiv:2503.10666v4 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have become widely used across various domains spanning search engines, code generation, and text creation. Howev

How we built the most performant DeepSeek V3.2, MiniMax-M2.5 and Qwen 3.5 397B on DigitalOcean Serverless Inference

Model ReleasesDGX agent

DigitalOcean describes their optimization and deployment of three large language models (DeepSeek V3.2, MiniMax-M2.5, and Qwen 3.5 397B) on their Serverless Inference platform, likely leveraging NVIDI

Isotonic Layer: A Unified Framework for Recommendation Calibration and Debiasing

SafetyDGX agent

arXiv:2603.06589v2 Announce Type: replace-cross Abstract: Model calibration and debiasing are fundamental yet operationally expensive challenges in large-scale recommendation systems. Existing approac

KLong: Training LLM Agent for Extremely Long-horizon Tasks

Model ReleasesDGX agent

arXiv:2602.17547v3 Announce Type: replace Abstract: This paper introduces KLong, an open-source LLM agent trained to solve extremely long-horizon tasks. The principle is to first cold-start the model

MEG-RAG: Quantifying Multi-modal Evidence Grounding for Evidence Selection in RAG

Model ReleasesDGX agent

arXiv:2604.24564v1 Announce Type: new Abstract: Multimodal Retrieval-Augmented Generation (MRAG) addresses key limitations of Multimodal Large Language Models (MLLMs), such as hallucination and outdat

MEMCoder: Multi-dimensional Evolving Memory for Private-Library-Oriented Code Generation

Model ReleasesDGX agent

arXiv:2604.24222v1 Announce Type: cross Abstract: Large Language Models (LLMs) excel at general code generation, but their performance drops sharply in enterprise settings that rely on internal privat

MetaErr: Towards Predicting Error Patterns in Deep Neural Networks

Model ReleasesDGX agent

arXiv:2604.23289v1 Announce Type: cross Abstract: Due to the unprecedented success of deep learning, it has become an integral component in several multimedia computing applications in todays world. U

Mobile-R1: Towards Interactive Capability for VLM-Based Mobile Agent via Systematic Training

Model ReleasesDGX agent

arXiv:2506.20332v4 Announce Type: replace Abstract: Vision-language model-based mobile agents have gained the ability to understand complex instructions and mobile screenshots, benefiting from reinfor

MuSS: A Large-Scale Dataset and Cinematic Narrative Benchmark for Multi-Shot Subject-to-Video Generation

Model ReleasesDGX agent

arXiv:2604.23789v1 Announce Type: new Abstract: While video foundation models excel at single-shot generation, real-world cinematic storytelling inherently relies on complex multi-shot sequencing. Fur

MVIGER: Multi-View Variational Integration of Complementary Knowledge for Generative Recommender

ApplicationsDGX agent

arXiv:2408.08686v4 Announce Type: replace-cross Abstract: Language Models (LMs) have been widely used in recommender systems to incorporate textual information of items into item IDs, leveraging their

No Test Cases, No Problem: Distillation-Driven Code Generation for Scientific Workflows

Model ReleasesDGX agent

arXiv:2604.23106v1 Announce Type: cross Abstract: Existing multi-agent Large Language Model (LLM) frameworks for code generation typically use execution feedback and improve iteratively using Input/Ou

Nvidia introduces Nemotron 3 Nano Omni with vision and speech for powerful agentic AI use

Model ReleasesDGX agent

Nvidia Corp. today launched a powerful reasoning artificial intelligence model that unifies text, vision and speech, capable of acting as the “brains” of faster, smarter agentic AI applications. Dubbe

Patterns vs. Patients: Evaluating LLMs against Mental Health Professionals on Personality Disorder Diagnosis through First-Person Narratives

Model ReleasesDGX agent

arXiv:2512.20298v2 Announce Type: replace-cross Abstract: Growing reliance on LLMs for psychiatric self-assessment raises questions about their ability to interpret qualitative patient narratives. Thi

PoseX: AI Defeats Physics Approaches on Protein-Ligand Cross Docking

Model ReleasesDGX agent

arXiv:2505.01700v3 Announce Type: replace Abstract: Existing protein-ligand docking studies typically focus on the self-docking scenario, which is less practical in real applications. Moreover, some s

Progressive Approximation in Deep Residual Networks: Theory and Validation

Model ReleasesDGX agent

arXiv:2604.24154v1 Announce Type: cross Abstract: The Universal Approximation Theorem (UAT) guarantees universal function approximation but does not explain how residual models distribute approximatio

Protecting the Trace: A Principled Black-Box Approach Against Distillation Attacks

SafetyDGX agent

arXiv:2604.23238v1 Announce Type: cross Abstract: Frontier models push the boundaries of what is learnable at extreme computational costs, yet distillation via sampling reasoning traces exposes closed

Quantifying Divergence in Inter-LLM Communication Through API Retrieval and Ranking

SafetyDGX agent

arXiv:2604.22760v1 Announce Type: cross Abstract: Large language models (LLMs) increasingly operate as autonomous agents that reason over external APIs to perform complex tasks. However, their reliabi

RealFin: How Well Do LLMs Reason About Finance When Users Leave Things Unsaid?

Model ReleasesDGX agent

arXiv:2602.07096v2 Announce Type: replace-cross Abstract: Reliable financial reasoning requires knowing not only how to answer, but also when an answer cannot be justified. In real financial practice,

Resource-Lean Lexicon Induction for German Dialects

Model ReleasesDGX agent

arXiv:2604.23824v1 Announce Type: new Abstract: Automatic induction of high-quality dictionaries is essential for building lexical resources, yet low-resource languages and dialects pose several chall

Self Knowledge Re-expression: A Fully Local Method for Adapting LLMs to Tasks Using Intrinsic Knowledge

Local AiDGX agent

arXiv:2604.22939v1 Announce Type: cross Abstract: While the next-token prediction (NTP) paradigm enables large language models (LLMs) to express their intrinsic knowledge, its sequential nature constr

SemiSAM-O1: How far can we push the boundary of annotation-efficient medical image segmentation?

ResearchDGX agent

arXiv:2604.24109v1 Announce Type: new Abstract: Semi-supervised learning (SSL) has become a promising solution to alleviate the annotation burden of deep learning-based medical image segmentation mode

SEVerA: Verified Synthesis of Self-Evolving Agents

Model ReleasesDGX agent

arXiv:2603.25111v2 Announce Type: replace Abstract: Recent advances have shown the effectiveness of self-evolving LLM agents on tasks such as program repair and scientific discovery. In this paradigm,

Skill Retrieval Augmentation for Agentic AI

Model ReleasesDGX agent

arXiv:2604.24594v1 Announce Type: cross Abstract: As large language models (LLMs) evolve into agentic problem solvers, they increasingly rely on external, reusable skills to handle tasks beyond their

Stabilizing Efficient Reasoning with Step-Level Advantage Selection

Model ReleasesDGX agent

arXiv:2604.24003v1 Announce Type: new Abstract: Large language models (LLMs) achieve strong reasoning performance by allocating substantial computation at inference time, often generating long and ver

Switch Attention: Towards Dynamic and Fine-grained Hybrid Transformers

Model ReleasesDGX agent

arXiv:2603.26380v2 Announce Type: replace Abstract: The attention mechanism has been the core component in modern transformer architectures. However, the computation of standard full attention scales

Symmetric Equilibrium Propagation for Thermodynamic Diffusion Training

Model ReleasesDGX agent

arXiv:2604.23806v1 Announce Type: cross Abstract: The reverse process in score-based diffusion models is formally equivalent to overdamped Langevin dynamics in a time-dependent energy landscape. In ou

SynthPert: Enhancing LLM Biological Reasoning via Synthetic Reasoning Traces for Cellular Perturbation Prediction

Model ReleasesDGX agent

arXiv:2509.25346v2 Announce Type: replace Abstract: Predicting cellular responses to genetic perturbations represents a fundamental challenge in systems biology, critical for advancing therapeutic dis

Test-Time Adaptation for Unsupervised Combinatorial Optimization

Local AiDGX agent

arXiv:2601.21048v2 Announce Type: replace Abstract: Unsupervised neural combinatorial optimization (NCO) enables learning powerful solvers without access to ground-truth solutions. Existing approaches

The new LLM trained only on pre-1931 text is small enough that it can potentially run on device, so, with the right tools, you can get a ful…

ApplicationsDGX agent

The new LLM trained only on pre-1931 text is small enough that it can potentially run on device, so, with the right tools, you can get a fully vintage version of Siri, but from the era of Downton Abbe

The Price of Agreement: Measuring LLM Sycophancy in Agentic Financial Applications

Model ReleasesDGX agent

arXiv:2604.24668v1 Announce Type: new Abstract: Given the increased use of LLMs in financial systems today, it becomes important to evaluate the safety and robustness of such systems. One failure mode

Try now: https://www.together.ai/models/nvidia-nemotron-3-nano-omni#

Model ReleasesDGX agent

NVIDIA Nemotron-3 Nano Omni is now available to try through Together AI's platform, offering access to a compact multimodal model capable of processing both text and audio inputs. This announcement hi

Welcome to the agentic era: Public sector highlights and reflections from Next ‘26

Model ReleasesDGX agent

Welcome to the agentic era! Last week, leaders from our public sector customer and partner ecosystem took the stage at Google Cloud Next to share how they are leveraging AI and agents to scale their i

When VLMs 'Fix' Students: Identifying and Penalizing Over-Correction in the Evaluation of Multi-line Handwritten Math OCR

Model ReleasesDGX agent

arXiv:2604.22774v1 Announce Type: cross Abstract: Accurate transcription of handwritten mathematics is crucial for educational AI systems, yet current benchmarks fail to evaluate this capability prope

World-R1: Reinforcing 3D Constraints for Text-to-Video Generation

SafetyDGX agent

arXiv:2604.24764v1 Announce Type: new Abstract: Recent video foundation models demonstrate impressive visual synthesis but frequently suffer from geometric inconsistencies. While existing methods atte

27 Apr 2026

Bridging the Long-Tail Gap: Robust Retrieval-Augmented Relation Completion via Multi-Stage Paraphrase Infusion

Model ReleasesDGX agent

arXiv:2604.22261v1 Announce Type: new Abstract: Large language models (LLMs) struggle with relation completion (RC), both with and without retrieval-augmented generation (RAG), particularly when the r

Call-Chain-Aware LLM-Based Test Generation for Java Projects

Model ReleasesDGX agent

arXiv:2604.22046v1 Announce Type: cross Abstract: Large language models (LLMs) have recently shown strong potential for generating project-level unit tests. However, existing state-of-the-art approach

Causal Concept Graphs in LLM Latent Space for Stepwise Reasoning

ResearchDGX agent

arXiv:2603.10377v2 Announce Type: replace-cross Abstract: Sparse autoencoders can localize where concepts live in language models, but not how they interact during multi-step reasoning. We propose Cau

DocPrune:Efficient Document Question Answering via Background, Question, and Comprehension-aware Token Pruning

ResearchDGX agent

arXiv:2604.22281v1 Announce Type: new Abstract: Recent advances in vision-language models have demonstrated remarkable performance across diverse multi-modal tasks, including document question answeri

Flux 2 dev

Local AiDGX agent

FLUX 2 Dev is an open-weight, 32-billion-parameter AI model developed by Black Forest Labs for text-to-image generation and advanced image editing . The model is available in ComfyUI and Diffusers fra

How Do AI Agents Spend Your Money? Analyzing and Predicting Token Consumption in Agentic Coding Tasks

Model ReleasesDGX agent

arXiv:2604.22750v1 Announce Type: new Abstract: The wide adoption of AI agents in complex human workflows is driving rapid growth in LLM token consumption. When agents are deployed on tasks that requi

Lately I've been having fun with running coding agents fully locally. The setup I landed on is: - Pi agent - Gemma 4 26B A4B - Server of cho…

Model ReleasesDGX agent

Lately I've been having fun with running coding agents fully locally. The setup I landed on is: - Pi agent - Gemma 4 26B A4B - Server of choice: LM Studio/Ollama/llama.cpp I wrote a step-by-step guide

Learning Coverage- and Power-Optimal Transmitter Placement from Building Maps: A Comparative Study of Direct and Indirect Neural Approaches

Model ReleasesDGX agent

arXiv:2604.22056v1 Announce Type: new Abstract: Optimal wireless transmitter placement is a central task in radio-network planning, yet exhaustive search becomes prohibitively expensive at scale. This

Multimodal Diffusion to Mutually Enhance Polarized Light and Low Resolution EBSD Data

TutorialsDGX agent

arXiv:2604.22212v1 Announce Type: cross Abstract: In spite of the utility of 3-D electron back-scattered diffraction (EBSD) microscopy, the data collection process can be time-consuming with serial-se

Optimal sequential decision-making for error propagation mitigation in digital twins

Model ReleasesDGX agent

arXiv:2604.22168v1 Announce Type: new Abstract: Here, we explore the problem of error propagation mitigation in modular digital twins as a sequential decision process. Building on a companion study th

Outcome Rewards Do Not Guarantee Verifiable or Causally Important Reasoning

ResearchDGX agent

arXiv:2604.22074v1 Announce Type: new Abstract: Reinforcement Learning from Verifiable Rewards (RLVR) on chain-of-thought reasoning has become a standard part of language model post-training recipes.

Rethinking Token Pruning for Historical Screenshots in GUI Visual Agents: Semantic, Spatial, and Temporal Perspectives

ResearchDGX agent

arXiv:2603.26041v3 Announce Type: replace Abstract: In recent years, GUI visual agents built upon Multimodal Large Language Models (MLLMs) have demonstrated strong potential in navigation tasks. Howev

← Previous
1…362363364365366…1050
Next →