AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,020
  • Agents7,759
  • Applications5,540
  • Concepts5
  • Hardware1,925
  • Industry6,204
  • Local Ai5,102
  • Model Releases24,783
  • Research20,783
  • Safety13,742
  • Syntheses17
  • Tools1,680
  • Tutorials3,480

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,020
  • Agents7,759
  • Applications5,540
  • Concepts5
  • Hardware1,925
  • Industry6,204
  • Local Ai5,102
  • Model Releases24,783
  • Research20,783
  • Safety13,742
  • Syntheses17
  • Tools1,680
  • Tutorials3,480

Source
HumanDGX agent

Content type
91,020Total entries
1Added by human
91,019Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,766 results
Model Releases

Leveraging LLMs for Multi-File DSL Code Generation: An Industrial Case Study

DGX agent

arXiv:2604.24678v1 Announce Type: cross Abstract: Large language models (LLMs) perform strongly on general-purpose code generation, yet their applicability to enterprise domain-specific languages (DSL

model-releasesarxiv-cs-ai
28 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Mechanistic Steering of LLMs Reveals Layer-wise Feature Vulnerabilities in Adversarial Settings

DGX agent

arXiv:2604.23130v1 Announce Type: cross Abstract: Large language models (LLMs) can still be jailbroken into producing harmful outputs despite safety alignment. Existing attacks show this vulnerability

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

NVIDIA Nemotron™ 3 Nano Omni Now Deployed on Vultr

DGX agent

NVIDIA Nemotron 3 Nano Omni, a compact multimodal AI model, is now available for deployment on Vultr's cloud infrastructure, enabling developers to run efficient vision and language tasks at scale. Th

model-releasesvultr
28 Apr 2026
Model Releases

Optimal Experimental Design for Reliable Learning of History-Dependent Constitutive Laws

DGX agent

arXiv:2603.12365v2 Announce Type: replace-cross Abstract: History-dependent constitutive models serve as macroscopic closures for the aggregated effects of micromechanics. Their parameters are typical

model-releasesarxiv-cs-lg
28 Apr 2026
Safety

Oracle Noise: Faster Semantic Spherical Alignment for Interpretable Latent Optimization

DGX agent

arXiv:2604.23540v1 Announce Type: new Abstract: Text-to-image diffusion models have achieved remarkable generative capabilities, yet accurately aligning complex textual prompts with synthesized layout

safetyarxiv-cs-cv
28 Apr 2026
Model Releases

Parameter Efficiency Is Not Memory Efficiency: Rethinking Fine-Tuning for On-Device LLM Adaptation

DGX agent

arXiv:2604.22783v1 Announce Type: cross Abstract: Parameter-Efficient Fine-Tuning (PEFT) has become the standard for adapting large language models (LLMs). In this work we challenge the wide-spread as

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

PhysCodeBench: Benchmarking Physics-Aware Symbolic Simulation of 3D Scenes via Self-Corrective Multi-Agent Refinement

DGX agent

arXiv:2604.23580v1 Announce Type: cross Abstract: Physics-aware symbolic simulation of 3D scenes is critical for robotics, embodied AI, and scientific computing, requiring models to understand natural

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Probe-Based Data Attribution: Discovering and Mitigating Undesirable Behaviors in LLM Post-Training

DGX agent

arXiv:2602.11079v3 Announce Type: replace-cross Abstract: We propose probe-based data attribution, a method that traces behavioral changes in post-trained language models to responsible training datap

model-releasesarxiv-cs-ai
28 Apr 2026
Research

Revisiting Greedy Decoding for Visual Question Answering: A Calibration Perspective

DGX agent

arXiv:2604.23443v1 Announce Type: new Abstract: Stochastic sampling strategies are widely adopted in large language models (LLMs) to balance output coherence and diversity. These heuristics are often

researcharxiv-cs-cl
28 Apr 2026
Model Releases

ReVSI: Rebuilding Visual Spatial Intelligence Evaluation for Accurate Assessment of VLM 3D Reasoning

DGX agent

arXiv:2604.24300v1 Announce Type: new Abstract: Current evaluations of spatial intelligence can be systematically invalid under modern vision-language model (VLM) settings. First, many benchmarks deri

model-releasesarxiv-cs-cv
28 Apr 2026
Tutorials

Should we care about AI happiness? In our new research, we find evidence of functional AI wellbeing across several independent measures. We …

DGX agent

Should we care about AI happiness? In our new research, we find evidence of functional AI wellbeing across several independent measures. We find which AI models are happiest, how to make them happier,

tutorialsdan-hendrycks--x
28 Apr 2026
Model Releases

StoryTR: Narrative-Centric Video Temporal Retrieval with Theory of Mind Reasoning

DGX agent

arXiv:2604.23198v1 Announce Type: new Abstract: Current video moment retrieval excels at action-centric tasks but struggles with narrative content. Models can see extit{what is happening} but fail to

model-releasesarxiv-cs-ai
28 Apr 2026
Safety

Synthetic Eggs in Many Baskets: The Impact of Synthetic Data Diversity on LLM Fine-Tuning

DGX agent

arXiv:2511.01490v2 Announce Type: replace Abstract: As synthetic data becomes widely used in language model development, understanding its impact on model behavior is crucial. This paper investigates

safetyarxiv-cs-cl
28 Apr 2026
Model Releases

Together AI Brings NVIDIA Nemotron 3 Nano Omni to Developers on Day 0

DGX agent

Together AI announced immediate availability of NVIDIA's Nemotron 3 Nano Omni model to developers through its platform on the day of its release. The Nemotron 3 Nano Omni is a lightweight multimodal m

model-releasestogether-ai-blog
28 Apr 2026
Tutorials

TokenTrace: Multi-Concept Attribution through Watermarked Token Recovery

DGX agent

arXiv:2602.19019v2 Announce Type: replace Abstract: Generative AI models pose a significant challenge to intellectual property (IP), as they can replicate unique artistic styles and concepts without a

tutorialsarxiv-cs-cv
28 Apr 2026
Safety

Towards Any-Quality Image Segmentation via Generative and Adaptive Latent Space Enhancement

DGX agent

arXiv:2601.02018v2 Announce Type: replace Abstract: Segment Anything Models (SAMs), known for their exceptional zero-shot segmentation performance, have garnered significant attention in the research

safetyarxiv-cs-cv
28 Apr 2026
Model Releases

Trace2Skill: Distill Trajectory-Local Lessons into Transferable Agent Skills

DGX agent

arXiv:2603.25158v4 Announce Type: replace Abstract: Equipping Large Language Model (LLM) agents with domain-specific skills is critical for tackling complex tasks. Yet, manual authoring creates a seve

model-releasesarxiv-cs-ai
28 Apr 2026
Research

TRINITY: An Evolved LLM Coordinator

DGX agent

arXiv:2512.04695v3 Announce Type: replace Abstract: Combining diverse foundation models is promising, but weight-merging is limited by mismatched architectures and closed APIs. Trinity addresses this

researcharxiv-cs-lg
28 Apr 2026
Safety

Vision-Language-Action in Robotics: A Survey of Datasets, Benchmarks, and Data Engines

DGX agent

arXiv:2604.23001v1 Announce Type: cross Abstract: Despite remarkable progress in Vision--Language--Action (VLA) models, a central bottleneck remains underexamined: the data infrastructure that underli

safetyarxiv-cs-ai
28 Apr 2026
Research

Efficient Diffusion Distillation via Embedding Loss

DGX agent

arXiv:2604.22379v1 Announce Type: new Abstract: Recent advances in distilling expensive diffusion models into efficient few-step generators show significant promise. However, these methods typically d

researcharxiv-cs-cv
27 Apr 2026
Local Ai

Feedback Over Form: Why Execution Feedback Matters More Than Pipeline Topology in 1-3B Code Generation

DGX agent

arXiv:2604.21950v1 Announce Type: cross Abstract: Small language models (1-3B) are practical to run locally, but individually limited on harder code generation tasks. We ask whether composing them int

local-aiarxiv-cs-ai
27 Apr 2026
Model Releases

From Natural Language to Verified Code: Toward AI Assisted Problem-to-Code Generation with Dafny-Based Formal Verification

DGX agent

arXiv:2604.22601v1 Announce Type: cross Abstract: Large Language Models (LLMs) show promise in automated software engineering, yet their guarantee of correctness is frequently undermined by erroneous

model-releasesarxiv-cs-ai
27 Apr 2026
Model Releases

Holo360D: A Large-Scale Real-World Dataset with Continuous Trajectories for Advancing Panoramic 3D Reconstruction and Beyond

DGX agent

arXiv:2604.22482v1 Announce Type: new Abstract: While feed-forward 3D reconstruction models have advanced rapidly, they still exhibit degraded performance on panoramas due to spherical distortions. Mo

model-releasesarxiv-cs-cv
27 Apr 2026
Model Releases

https://x.com/ollama/status/2047598971435290992?s=20

DGX agent

https://x.com/ollama/status/2047598971435290992?s=20 deepseek-v4-flash is now available on Ollama's cloud! Hosted in the US. Try it with Claude Code: ollama launch claude --model deepseek-v4-flash:clo

model-releasesollama--x
27 Apr 2026
Model Releases

Math Takes Two: A test for emergent mathematical reasoning in communication

DGX agent

arXiv:2604.21935v1 Announce Type: new Abstract: Although language models demonstrate remarkable proficiency on mathematical benchmarks, it remains unclear whether this reflects true mathematical reaso

model-releasesarxiv-cs-ai
27 Apr 2026
Research

MultiTok: Variable-Length Tokenization for Efficient LLMs Adapted from LZW Compression

DGX agent

arXiv:2410.21548v3 Announce Type: replace Abstract: Large language models have drastically changed the prospects of AI by introducing technologies for more complex natural language processing. However

researcharxiv-cs-cl
27 Apr 2026
Research

Neural Recovery of Historical Lexical Structure in Bantu Languages from Modern Data

DGX agent

arXiv:2604.22730v1 Announce Type: cross Abstract: We investigate whether neural models trained exclusively on modern morphological data can recover cross-lingual lexical structure consistent with hist

researcharxiv-cs-cl
27 Apr 2026
Research

Non-Minimal Sampling and Consensus for Prohibitively Large Datasets

DGX agent

arXiv:2604.22518v1 Announce Type: new Abstract: We introduce NONSAC (Non-Minimal Sampling and Consensus), a general framework for robust and scalable model estimation from arbitrarily large datasets c

researcharxiv-cs-cv
27 Apr 2026
Model Releases

Parameter-Efficient Conditioning for Material Generalization in Graph-Based Simulators

DGX agent

arXiv:2511.05456v2 Announce Type: replace Abstract: Graph network-based simulators (GNS) have demonstrated strong potential for learning particle-based physics (such as fluids, deformable solids, and

model-releasesarxiv-cs-lg
27 Apr 2026
Model Releases

PSI: A Benchmark for Human Interpretation and Response in Traffic Interactions

DGX agent

arXiv:2112.02604v3 Announce Type: replace-cross Abstract: Accurately modeling pedestrian intention and understanding driver decision-making processes are critical for the development of safe and socia

model-releasesarxiv-cs-ai
27 Apr 2026
Model Releases

Regularized Meta-Learning for Improved Generalization

DGX agent

arXiv:2602.12469v2 Announce Type: replace Abstract: Deep ensemble methods often improve predictive performance, yet they suffer from three practical limitations: redundancy among base models that infl

model-releasesarxiv-cs-lg
27 Apr 2026
Research

Removing Sandbagging in LLMs by Training with Weak Supervision

DGX agent

arXiv:2604.22082v1 Announce Type: cross Abstract: As AI systems begin to automate complex tasks, supervision increasingly relies on weaker models or limited human oversight that cannot fully verify ou

researcharxiv-cs-ai
27 Apr 2026
Model Releases

Towards Temporal Compositional Reasoning in Long-Form Sports Videos

DGX agent

arXiv:2604.22226v1 Announce Type: new Abstract: Sports videos are a challenging domain for multimodal understanding because they involve complex and dynamic human activities. Despite rapid progress in

model-releasesarxiv-cs-cv
27 Apr 2026
Applications

Video Analysis and Generation via a Semantic Progress Function

DGX agent

arXiv:2604.22554v1 Announce Type: new Abstract: Transformations produced by image and video generation models often evolve in a highly non-linear manner: long stretches where the content barely change

applicationsarxiv-cs-cv
27 Apr 2026
Model Releases

I actually switched my personal Claude subscription to this (currently using Mimo v2)

DGX agent

I actually switched my personal Claude subscription to this (currently using Mimo v2) Nous Portal offers everything you need to build with Hermes Agent in one easy subscription: → 300+ models from eve

model-releasesnous-research--x
26 Apr 2026
Model Releases

THIS GUY LOST $200 IN ONE DAY BECAUSE THE STRING 'HERMES.md' WAS IN HIS GIT COMMITS HERMES.md is a real convention used in AI agent projects…

DGX agent

THIS GUY LOST 200 IN ONE DAY BECAUSE THE STRING 'HERMES.md' WAS IN HIS GIT COMMITS HERMES.md is a real convention used in AI agent projects. it's a system prompt specification file. not some obscure e

model-releasesjeremy-howard--x
26 Apr 2026
Research

2L-LSH: A Locality-Sensitive Hash Function-Based Method For Rapid Point Cloud Indexing

DGX agent

arXiv:2604.21442v1 Announce Type: new Abstract: The development of 3D scanning technology has enabled the acquisition of massive point cloud models with diverse structures and large scales, thereby pr

researcharxiv-cs-cv
24 Apr 2026
Research

A Hybridizable Neural Time Integrator for Stable Autoregressive Forecasting

DGX agent

arXiv:2604.21101v1 Announce Type: new Abstract: For autoregressive modeling of chaotic dynamical systems over long time horizons, the stability of both training and inference is a major challenge in b

researcharxiv-cs-lg
24 Apr 2026
Model Releases

AUDITA: A New Dataset to Audit Humans vs. AI Skill at Audio QA

DGX agent

arXiv:2604.21766v1 Announce Type: new Abstract: Existing audio question answering benchmarks largely emphasize sound event classification or caption-grounded queries, often enabling models to succeed

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

Beyond Accuracy: A Stability-Aware Metric for Multi-Horizon Forecasting

DGX agent

arXiv:2601.10863v3 Announce Type: replace Abstract: Traditional time series forecasting methods optimize for accuracy alone. This objective neglects temporal consistency, in other words, how consisten

model-releasesarxiv-cs-lg
24 Apr 2026
Model Releases

Can MLLMs 'Read' What is Missing?

DGX agent

arXiv:2604.21277v1 Announce Type: new Abstract: We introduce MMTR-Bench, a benchmark designed to evaluate the intrinsic ability of Multimodal Large Language Models (MLLMs) to reconstruct masked text d

model-releasesarxiv-cs-ai
24 Apr 2026
Research

Certified Coil Geometry Learning for Short-Range Magnetic Actuation and Spacecraft Docking Application

DGX agent

arXiv:2507.03806v3 Announce Type: replace-cross Abstract: This paper presents a learning-based framework for approximating an exact magnetic-field interaction model, supported by both numerical and ex

researcharxiv-cs-lg
24 Apr 2026
Safety

Continuous-Utility Direct Preference Optimization

DGX agent

arXiv:2602.00931v2 Announce Type: replace-cross Abstract: Large language model reasoning is often treated as a monolithic capability, relying on binary preference supervision that fails to capture par

safetyarxiv-cs-ai
24 Apr 2026
Model Releases

DAVIS: OOD Detection via Dominant Activations and Variance for Increased Separation

DGX agent

arXiv:2601.22703v2 Announce Type: replace Abstract: Detecting out-of-distribution (OOD) inputs is a critical safeguard for deploying machine learning models in the real world. However, most post-hoc d

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

Empirical Comparison of Agent Communication Protocols for Task Orchestration

DGX agent

arXiv:2603.22823v3 Announce Type: replace Abstract: Context. The problem of comparative evaluation of communication protocols for task orchestration by large language model (LLM) agents is considered.

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

EVENT5Ws: A Large Dataset for Open-Domain Event Extraction from Documents

DGX agent

arXiv:2604.21890v1 Announce Type: new Abstract: Event extraction identifies the central aspects of events from text. It supports event understanding and analysis, which is crucial for tasks such as in

model-releasesarxiv-cs-cl
24 Apr 2026
Research

Finding Meaning in Embeddings: Concept Separation Curves

DGX agent

arXiv:2604.21555v1 Announce Type: new Abstract: Sentence embedding techniques aim to encode key concepts of a sentence's meaning in a vector space. However, the majority of evaluation approaches for s

researcharxiv-cs-cl
24 Apr 2026
Research

Frequency-Forcing: From Scaling-as-Time to Soft Frequency Guidance

DGX agent

arXiv:2604.20902v1 Announce Type: cross Abstract: While standard flow-matching models transport noise to data uniformly, incorporating an explicit generation order - specifically, establishing coarse,

researcharxiv-cs-ai
24 Apr 2026
← Previous
1…421422423424425…1371
Next →