AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,429
  • Agents7,559
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,170
  • Local Ai4,935
  • Model Releases23,918
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,429
  • Agents7,559
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,170
  • Local Ai4,935
  • Model Releases23,918
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlog
88,429Total entries
1Added by human
88,428Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,648 results
Applications

RELIC: Evaluating Complex Reasoning via the Recognition of Languages In-Context

DGX agent

arXiv:2506.05205v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly used to solve complex tasks where they must retrieve and compose many pieces of in-context information

applicationsarxiv-cs-cl
29 Apr 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

SIEVES: Selective Prediction Generalizes through Visual Evidence Scoring

DGX agent

arXiv:2604.25855v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) achieve ever-stronger performance on visual-language tasks. Even as traditional visual question answering bench

model-releasesarxiv-cs-cv
29 Apr 2026
Model Releases

SynMotion: Semantic-Visual Adaptation for Motion Customized Video Generation

DGX agent

arXiv:2506.23690v2 Announce Type: replace Abstract: Diffusion-based video motion customization facilitates the acquisition of human motion representations from a few video samples, while achieving arb

model-releasesarxiv-cs-cv
29 Apr 2026
Model Releases

The founder’s AI foundation: The top announcements for startups from Next ‘26

DGX agent

The momentum is undeniable: the world’s fastest-growing AI startups are building with Google Cloud. Instead of stitching together fragmented point solutions, founders are building their businesses her

model-releasesgoogle-cloud-ai
29 Apr 2026
Model Releases

Towards Unified Multi-task EEG Analysis with Low-Rank Adaptation

DGX agent

arXiv:2604.25131v1 Announce Type: new Abstract: Recent self-supervised pre-training methods for electroencephalogram (EEG) have shown promising results. However, the pre-trained models typically requi

model-releasesarxiv-cs-lg
29 Apr 2026
Tutorials

2D Pre-Training for 3D Pose Estimation

DGX agent

arXiv:2604.22830v1 Announce Type: new Abstract: Pre-training is a general method that is used in a range of deep learning tasks. By first training a model on one task, and then further training on the

tutorialsarxiv-cs-cv
28 Apr 2026
Model Releases

A BERTology View of LLM Orchestrations: Token- and Layer-Selective Probes for Efficient Single-Pass Classification

DGX agent

arXiv:2601.13288v2 Announce Type: replace Abstract: Production LLM systems often rely on separate models for safety and other classification-heavy steps, increasing latency, VRAM footprint, and operat

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

A Parametric Memory Head for Continual Generative Retrieval

DGX agent

arXiv:2604.23388v1 Announce Type: cross Abstract: Generative information retrieval (GenIR) consolidates retrieval into a single neural model that decodes document identifiers (docids) directly from qu

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Accelerating Quantum Materials Characterization: Hybrid Active Learning for Autonomous Spin Wave Spectroscopy

DGX agent

arXiv:2604.23821v1 Announce Type: cross Abstract: Autonomous neutron spectroscopy must solve three distinct tasks: detection (where is the signal?), inference (which Hamiltonian governs it?), and refi

model-releasesarxiv-cs-lg
28 Apr 2026
Model Releases

All That Glitters Is Not Audio: Rethinking Text Priors and Audio Reliance in Audio-Language Evaluation

DGX agent

arXiv:2604.24401v1 Announce Type: cross Abstract: Large Audio-Language Models show consistent performance gains across speech and audio benchmarks, yet high scores may not reflect true auditory percep

model-releasesarxiv-cs-ai
28 Apr 2026
Research

Always Tell Me The Odds: Fine-grained Conditional Probability Estimation

DGX agent

arXiv:2505.01595v2 Announce Type: replace-cross Abstract: We present a state-of-the-art model for fine-grained probability estimation of propositions conditioned on context. Recent advances in large l

researcharxiv-cs-ai
28 Apr 2026
Model Releases

Applications of the Transformer Architecture in AI-Assisted English Reading Comprehension

DGX agent

arXiv:2604.23615v1 Announce Type: cross Abstract: This paper studies interpretable and fair artificial intelligence architectures for understanding English reading. Introduced transformer-based models

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

AutoGUI-v2: A Comprehensive Multi-Modal GUI Functionality Understanding Benchmark

DGX agent

arXiv:2604.24441v1 Announce Type: new Abstract: Autonomous agents capable of navigating Graphical User Interfaces (GUIs) hold the potential to revolutionize digital productivity. However, achieving tr

model-releasesarxiv-cs-cv
28 Apr 2026
Tutorials

AV-Master: Dual-Path Comprehensive Perception Makes Better Audio-Visual Question Answering

DGX agent

arXiv:2510.18346v2 Announce Type: replace Abstract: Audio-Visual Question Answering (AVQA) requires models to effectively utilize both visual and auditory modalities to answer complex and diverse ques

tutorialsarxiv-cs-cv
28 Apr 2026
Model Releases

Benchmarking Testing in Automated Theorem Proving

DGX agent

arXiv:2604.23698v1 Announce Type: new Abstract: Recent advances in large language models (LLMs) have shown promise in formal theorem proving, yet evaluating semantic correctness remains challenging. E

model-releasesarxiv-cs-cl
28 Apr 2026
Research

BrickNet: Graph-Backed Generative Brick Assembly

DGX agent

arXiv:2604.22984v1 Announce Type: new Abstract: We train a language model to generate LEGO-brick build sequences. While prior work has been restricted to discrete, voxel-like towers, we consider a muc

researcharxiv-cs-cv
28 Apr 2026
Model Releases

Complexity of Linear Regions in Self-supervised Deep ReLU Networks

DGX agent

arXiv:2604.24393v1 Announce Type: cross Abstract: There has been growing interest in studying the complexity of Rectified Linear Unit (ReLU) based activation networks. Recent work investigates the evo

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

CRISP: Persistent Concept Unlearning via Sparse Autoencoders

DGX agent

arXiv:2508.13650v3 Announce Type: replace Abstract: As large language models (LLMs) are increasingly deployed in real-world applications, the need to selectively remove unwanted knowledge while preser

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

Defective Task Descriptions in LLM-Based Code Generation: Detection and Analysis

DGX agent

arXiv:2604.24703v1 Announce Type: cross Abstract: Large language models are widely used for code generation, yet they rely on an implicit assumption that the task descriptions are sufficiently detaile

model-releasesarxiv-cs-ai
28 Apr 2026
Tutorials

Doloris: Dual Conditional Diffusion Implicit Bridges with Sparsity Masking Strategy for Unpaired Single-Cell Perturbation Estimation

DGX agent

arXiv:2506.21107v3 Announce Type: replace Abstract: Estimating single-cell responses across various perturbations facilitates the identification of key genes and enhances drug screening, significantly

tutorialsarxiv-cs-lg
28 Apr 2026
Model Releases

EPM-RL: Reinforcement Learning for On-Premise Product Mapping in E-Commerce

DGX agent

arXiv:2604.23993v1 Announce Type: cross Abstract: Product mapping, the task of deciding whether two e-commerce listings refer to the same product, is a core problem for price monitoring and channel vi

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Global Context or Local Detail? Adaptive Visual Grounding for Hallucination Mitigation

DGX agent

arXiv:2604.24396v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) are frequently undermined by object hallucination--generating content that contradicts visual reality--due to an over-re

model-releasesarxiv-cs-ai
28 Apr 2026
Local Ai

Kwai Summary Attention Technical Report

DGX agent

arXiv:2604.24432v1 Announce Type: cross Abstract: Long-context ability, has become one of the most important iteration direction of next-generation Large Language Models, particularly in semantic unde

local-aiarxiv-cs-ai
28 Apr 2026
Model Releases

Leveraging LLMs for Multi-File DSL Code Generation: An Industrial Case Study

DGX agent

arXiv:2604.24678v1 Announce Type: cross Abstract: Large language models (LLMs) perform strongly on general-purpose code generation, yet their applicability to enterprise domain-specific languages (DSL

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Mechanistic Steering of LLMs Reveals Layer-wise Feature Vulnerabilities in Adversarial Settings

DGX agent

arXiv:2604.23130v1 Announce Type: cross Abstract: Large language models (LLMs) can still be jailbroken into producing harmful outputs despite safety alignment. Existing attacks show this vulnerability

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

NVIDIA Nemotron™ 3 Nano Omni Now Deployed on Vultr

DGX agent

NVIDIA Nemotron 3 Nano Omni, a compact multimodal AI model, is now available for deployment on Vultr's cloud infrastructure, enabling developers to run efficient vision and language tasks at scale. Th

model-releasesvultr
28 Apr 2026
Model Releases

Optimal Experimental Design for Reliable Learning of History-Dependent Constitutive Laws

DGX agent

arXiv:2603.12365v2 Announce Type: replace-cross Abstract: History-dependent constitutive models serve as macroscopic closures for the aggregated effects of micromechanics. Their parameters are typical

model-releasesarxiv-cs-lg
28 Apr 2026
Safety

Oracle Noise: Faster Semantic Spherical Alignment for Interpretable Latent Optimization

DGX agent

arXiv:2604.23540v1 Announce Type: new Abstract: Text-to-image diffusion models have achieved remarkable generative capabilities, yet accurately aligning complex textual prompts with synthesized layout

safetyarxiv-cs-cv
28 Apr 2026
Model Releases

Parameter Efficiency Is Not Memory Efficiency: Rethinking Fine-Tuning for On-Device LLM Adaptation

DGX agent

arXiv:2604.22783v1 Announce Type: cross Abstract: Parameter-Efficient Fine-Tuning (PEFT) has become the standard for adapting large language models (LLMs). In this work we challenge the wide-spread as

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

PhysCodeBench: Benchmarking Physics-Aware Symbolic Simulation of 3D Scenes via Self-Corrective Multi-Agent Refinement

DGX agent

arXiv:2604.23580v1 Announce Type: cross Abstract: Physics-aware symbolic simulation of 3D scenes is critical for robotics, embodied AI, and scientific computing, requiring models to understand natural

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Probe-Based Data Attribution: Discovering and Mitigating Undesirable Behaviors in LLM Post-Training

DGX agent

arXiv:2602.11079v3 Announce Type: replace-cross Abstract: We propose probe-based data attribution, a method that traces behavioral changes in post-trained language models to responsible training datap

model-releasesarxiv-cs-ai
28 Apr 2026
Research

Revisiting Greedy Decoding for Visual Question Answering: A Calibration Perspective

DGX agent

arXiv:2604.23443v1 Announce Type: new Abstract: Stochastic sampling strategies are widely adopted in large language models (LLMs) to balance output coherence and diversity. These heuristics are often

researcharxiv-cs-cl
28 Apr 2026
Model Releases

ReVSI: Rebuilding Visual Spatial Intelligence Evaluation for Accurate Assessment of VLM 3D Reasoning

DGX agent

arXiv:2604.24300v1 Announce Type: new Abstract: Current evaluations of spatial intelligence can be systematically invalid under modern vision-language model (VLM) settings. First, many benchmarks deri

model-releasesarxiv-cs-cv
28 Apr 2026
Tutorials

Should we care about AI happiness? In our new research, we find evidence of functional AI wellbeing across several independent measures. We …

DGX agent

Should we care about AI happiness? In our new research, we find evidence of functional AI wellbeing across several independent measures. We find which AI models are happiest, how to make them happier,

tutorialsdan-hendrycks--x
28 Apr 2026
Model Releases

StoryTR: Narrative-Centric Video Temporal Retrieval with Theory of Mind Reasoning

DGX agent

arXiv:2604.23198v1 Announce Type: new Abstract: Current video moment retrieval excels at action-centric tasks but struggles with narrative content. Models can see extit{what is happening} but fail to

model-releasesarxiv-cs-ai
28 Apr 2026
Safety

Synthetic Eggs in Many Baskets: The Impact of Synthetic Data Diversity on LLM Fine-Tuning

DGX agent

arXiv:2511.01490v2 Announce Type: replace Abstract: As synthetic data becomes widely used in language model development, understanding its impact on model behavior is crucial. This paper investigates

safetyarxiv-cs-cl
28 Apr 2026
Model Releases

Together AI Brings NVIDIA Nemotron 3 Nano Omni to Developers on Day 0

DGX agent

Together AI announced immediate availability of NVIDIA's Nemotron 3 Nano Omni model to developers through its platform on the day of its release. The Nemotron 3 Nano Omni is a lightweight multimodal m

model-releasestogether-ai-blog
28 Apr 2026
Tutorials

TokenTrace: Multi-Concept Attribution through Watermarked Token Recovery

DGX agent

arXiv:2602.19019v2 Announce Type: replace Abstract: Generative AI models pose a significant challenge to intellectual property (IP), as they can replicate unique artistic styles and concepts without a

tutorialsarxiv-cs-cv
28 Apr 2026
Safety

Towards Any-Quality Image Segmentation via Generative and Adaptive Latent Space Enhancement

DGX agent

arXiv:2601.02018v2 Announce Type: replace Abstract: Segment Anything Models (SAMs), known for their exceptional zero-shot segmentation performance, have garnered significant attention in the research

safetyarxiv-cs-cv
28 Apr 2026
Model Releases

Trace2Skill: Distill Trajectory-Local Lessons into Transferable Agent Skills

DGX agent

arXiv:2603.25158v4 Announce Type: replace Abstract: Equipping Large Language Model (LLM) agents with domain-specific skills is critical for tackling complex tasks. Yet, manual authoring creates a seve

model-releasesarxiv-cs-ai
28 Apr 2026
Research

TRINITY: An Evolved LLM Coordinator

DGX agent

arXiv:2512.04695v3 Announce Type: replace Abstract: Combining diverse foundation models is promising, but weight-merging is limited by mismatched architectures and closed APIs. Trinity addresses this

researcharxiv-cs-lg
28 Apr 2026
Safety

Vision-Language-Action in Robotics: A Survey of Datasets, Benchmarks, and Data Engines

DGX agent

arXiv:2604.23001v1 Announce Type: cross Abstract: Despite remarkable progress in Vision--Language--Action (VLA) models, a central bottleneck remains underexamined: the data infrastructure that underli

safetyarxiv-cs-ai
28 Apr 2026
Research

Efficient Diffusion Distillation via Embedding Loss

DGX agent

arXiv:2604.22379v1 Announce Type: new Abstract: Recent advances in distilling expensive diffusion models into efficient few-step generators show significant promise. However, these methods typically d

researcharxiv-cs-cv
27 Apr 2026
Local Ai

Feedback Over Form: Why Execution Feedback Matters More Than Pipeline Topology in 1-3B Code Generation

DGX agent

arXiv:2604.21950v1 Announce Type: cross Abstract: Small language models (1-3B) are practical to run locally, but individually limited on harder code generation tasks. We ask whether composing them int

local-aiarxiv-cs-ai
27 Apr 2026
Model Releases

From Natural Language to Verified Code: Toward AI Assisted Problem-to-Code Generation with Dafny-Based Formal Verification

DGX agent

arXiv:2604.22601v1 Announce Type: cross Abstract: Large Language Models (LLMs) show promise in automated software engineering, yet their guarantee of correctness is frequently undermined by erroneous

model-releasesarxiv-cs-ai
27 Apr 2026
Model Releases

Holo360D: A Large-Scale Real-World Dataset with Continuous Trajectories for Advancing Panoramic 3D Reconstruction and Beyond

DGX agent

arXiv:2604.22482v1 Announce Type: new Abstract: While feed-forward 3D reconstruction models have advanced rapidly, they still exhibit degraded performance on panoramas due to spherical distortions. Mo

model-releasesarxiv-cs-cv
27 Apr 2026
Model Releases

https://x.com/ollama/status/2047598971435290992?s=20

DGX agent

https://x.com/ollama/status/2047598971435290992?s=20 deepseek-v4-flash is now available on Ollama's cloud! Hosted in the US. Try it with Claude Code: ollama launch claude --model deepseek-v4-flash:clo

model-releasesollama--x
27 Apr 2026
Model Releases

Math Takes Two: A test for emergent mathematical reasoning in communication

DGX agent

arXiv:2604.21935v1 Announce Type: new Abstract: Although language models demonstrate remarkable proficiency on mathematical benchmarks, it remains unclear whether this reflects true mathematical reaso

model-releasesarxiv-cs-ai
27 Apr 2026
← Previous
1…406407408409410…1326
Next →