AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,965
  • Agents7,446
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,131
  • Local Ai4,857
  • Model Releases23,360
  • Research19,834
  • Safety13,174
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,965
  • Agents7,446
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,131
  • Local Ai4,857
  • Model Releases23,360
  • Research19,834
  • Safety13,174
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

86,965Total entries
1Added by human
86,964Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,458 results
28 Apr 2026

Complexity of Linear Regions in Self-supervised Deep ReLU Networks

Model ReleasesDGX agent

arXiv:2604.24393v1 Announce Type: cross Abstract: There has been growing interest in studying the complexity of Rectified Linear Unit (ReLU) based activation networks. Recent work investigates the evo

CRISP: Persistent Concept Unlearning via Sparse Autoencoders

Model ReleasesDGX agent

arXiv:2508.13650v3 Announce Type: replace Abstract: As large language models (LLMs) are increasingly deployed in real-world applications, the need to selectively remove unwanted knowledge while preser

Defective Task Descriptions in LLM-Based Code Generation: Detection and Analysis

Model ReleasesDGX agent

arXiv:2604.24703v1 Announce Type: cross Abstract: Large language models are widely used for code generation, yet they rely on an implicit assumption that the task descriptions are sufficiently detaile

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Doloris: Dual Conditional Diffusion Implicit Bridges with Sparsity Masking Strategy for Unpaired Single-Cell Perturbation Estimation

TutorialsDGX agent

arXiv:2506.21107v3 Announce Type: replace Abstract: Estimating single-cell responses across various perturbations facilitates the identification of key genes and enhances drug screening, significantly

EPM-RL: Reinforcement Learning for On-Premise Product Mapping in E-Commerce

Model ReleasesDGX agent

arXiv:2604.23993v1 Announce Type: cross Abstract: Product mapping, the task of deciding whether two e-commerce listings refer to the same product, is a core problem for price monitoring and channel vi

Global Context or Local Detail? Adaptive Visual Grounding for Hallucination Mitigation

Model ReleasesDGX agent

arXiv:2604.24396v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) are frequently undermined by object hallucination--generating content that contradicts visual reality--due to an over-re

Kwai Summary Attention Technical Report

Local AiDGX agent

arXiv:2604.24432v1 Announce Type: cross Abstract: Long-context ability, has become one of the most important iteration direction of next-generation Large Language Models, particularly in semantic unde

Leveraging LLMs for Multi-File DSL Code Generation: An Industrial Case Study

Model ReleasesDGX agent

arXiv:2604.24678v1 Announce Type: cross Abstract: Large language models (LLMs) perform strongly on general-purpose code generation, yet their applicability to enterprise domain-specific languages (DSL

Mechanistic Steering of LLMs Reveals Layer-wise Feature Vulnerabilities in Adversarial Settings

Model ReleasesDGX agent

arXiv:2604.23130v1 Announce Type: cross Abstract: Large language models (LLMs) can still be jailbroken into producing harmful outputs despite safety alignment. Existing attacks show this vulnerability

NVIDIA Nemotron™ 3 Nano Omni Now Deployed on Vultr

Model ReleasesDGX agent

NVIDIA Nemotron 3 Nano Omni, a compact multimodal AI model, is now available for deployment on Vultr's cloud infrastructure, enabling developers to run efficient vision and language tasks at scale. Th

Optimal Experimental Design for Reliable Learning of History-Dependent Constitutive Laws

Model ReleasesDGX agent

arXiv:2603.12365v2 Announce Type: replace-cross Abstract: History-dependent constitutive models serve as macroscopic closures for the aggregated effects of micromechanics. Their parameters are typical

Oracle Noise: Faster Semantic Spherical Alignment for Interpretable Latent Optimization

SafetyDGX agent

arXiv:2604.23540v1 Announce Type: new Abstract: Text-to-image diffusion models have achieved remarkable generative capabilities, yet accurately aligning complex textual prompts with synthesized layout

Parameter Efficiency Is Not Memory Efficiency: Rethinking Fine-Tuning for On-Device LLM Adaptation

Model ReleasesDGX agent

arXiv:2604.22783v1 Announce Type: cross Abstract: Parameter-Efficient Fine-Tuning (PEFT) has become the standard for adapting large language models (LLMs). In this work we challenge the wide-spread as

PhysCodeBench: Benchmarking Physics-Aware Symbolic Simulation of 3D Scenes via Self-Corrective Multi-Agent Refinement

Model ReleasesDGX agent

arXiv:2604.23580v1 Announce Type: cross Abstract: Physics-aware symbolic simulation of 3D scenes is critical for robotics, embodied AI, and scientific computing, requiring models to understand natural

Probe-Based Data Attribution: Discovering and Mitigating Undesirable Behaviors in LLM Post-Training

Model ReleasesDGX agent

arXiv:2602.11079v3 Announce Type: replace-cross Abstract: We propose probe-based data attribution, a method that traces behavioral changes in post-trained language models to responsible training datap

Revisiting Greedy Decoding for Visual Question Answering: A Calibration Perspective

ResearchDGX agent

arXiv:2604.23443v1 Announce Type: new Abstract: Stochastic sampling strategies are widely adopted in large language models (LLMs) to balance output coherence and diversity. These heuristics are often

ReVSI: Rebuilding Visual Spatial Intelligence Evaluation for Accurate Assessment of VLM 3D Reasoning

Model ReleasesDGX agent

arXiv:2604.24300v1 Announce Type: new Abstract: Current evaluations of spatial intelligence can be systematically invalid under modern vision-language model (VLM) settings. First, many benchmarks deri

Should we care about AI happiness? In our new research, we find evidence of functional AI wellbeing across several independent measures. We …

TutorialsDGX agent

Should we care about AI happiness? In our new research, we find evidence of functional AI wellbeing across several independent measures. We find which AI models are happiest, how to make them happier,

StoryTR: Narrative-Centric Video Temporal Retrieval with Theory of Mind Reasoning

Model ReleasesDGX agent

arXiv:2604.23198v1 Announce Type: new Abstract: Current video moment retrieval excels at action-centric tasks but struggles with narrative content. Models can see extit{what is happening} but fail to

Synthetic Eggs in Many Baskets: The Impact of Synthetic Data Diversity on LLM Fine-Tuning

SafetyDGX agent

arXiv:2511.01490v2 Announce Type: replace Abstract: As synthetic data becomes widely used in language model development, understanding its impact on model behavior is crucial. This paper investigates

Together AI Brings NVIDIA Nemotron 3 Nano Omni to Developers on Day 0

Model ReleasesDGX agent

Together AI announced immediate availability of NVIDIA's Nemotron 3 Nano Omni model to developers through its platform on the day of its release. The Nemotron 3 Nano Omni is a lightweight multimodal m

TokenTrace: Multi-Concept Attribution through Watermarked Token Recovery

TutorialsDGX agent

arXiv:2602.19019v2 Announce Type: replace Abstract: Generative AI models pose a significant challenge to intellectual property (IP), as they can replicate unique artistic styles and concepts without a

Towards Any-Quality Image Segmentation via Generative and Adaptive Latent Space Enhancement

SafetyDGX agent

arXiv:2601.02018v2 Announce Type: replace Abstract: Segment Anything Models (SAMs), known for their exceptional zero-shot segmentation performance, have garnered significant attention in the research

Trace2Skill: Distill Trajectory-Local Lessons into Transferable Agent Skills

Model ReleasesDGX agent

arXiv:2603.25158v4 Announce Type: replace Abstract: Equipping Large Language Model (LLM) agents with domain-specific skills is critical for tackling complex tasks. Yet, manual authoring creates a seve

TRINITY: An Evolved LLM Coordinator

ResearchDGX agent

arXiv:2512.04695v3 Announce Type: replace Abstract: Combining diverse foundation models is promising, but weight-merging is limited by mismatched architectures and closed APIs. Trinity addresses this

Vision-Language-Action in Robotics: A Survey of Datasets, Benchmarks, and Data Engines

SafetyDGX agent

arXiv:2604.23001v1 Announce Type: cross Abstract: Despite remarkable progress in Vision--Language--Action (VLA) models, a central bottleneck remains underexamined: the data infrastructure that underli

27 Apr 2026

Efficient Diffusion Distillation via Embedding Loss

ResearchDGX agent

arXiv:2604.22379v1 Announce Type: new Abstract: Recent advances in distilling expensive diffusion models into efficient few-step generators show significant promise. However, these methods typically d

Feedback Over Form: Why Execution Feedback Matters More Than Pipeline Topology in 1-3B Code Generation

Local AiDGX agent

arXiv:2604.21950v1 Announce Type: cross Abstract: Small language models (1-3B) are practical to run locally, but individually limited on harder code generation tasks. We ask whether composing them int

From Natural Language to Verified Code: Toward AI Assisted Problem-to-Code Generation with Dafny-Based Formal Verification

Model ReleasesDGX agent

arXiv:2604.22601v1 Announce Type: cross Abstract: Large Language Models (LLMs) show promise in automated software engineering, yet their guarantee of correctness is frequently undermined by erroneous

Holo360D: A Large-Scale Real-World Dataset with Continuous Trajectories for Advancing Panoramic 3D Reconstruction and Beyond

Model ReleasesDGX agent

arXiv:2604.22482v1 Announce Type: new Abstract: While feed-forward 3D reconstruction models have advanced rapidly, they still exhibit degraded performance on panoramas due to spherical distortions. Mo

https://x.com/ollama/status/2047598971435290992?s=20

Model ReleasesDGX agent

https://x.com/ollama/status/2047598971435290992?s=20 deepseek-v4-flash is now available on Ollama's cloud! Hosted in the US. Try it with Claude Code: ollama launch claude --model deepseek-v4-flash:clo

Math Takes Two: A test for emergent mathematical reasoning in communication

Model ReleasesDGX agent

arXiv:2604.21935v1 Announce Type: new Abstract: Although language models demonstrate remarkable proficiency on mathematical benchmarks, it remains unclear whether this reflects true mathematical reaso

MultiTok: Variable-Length Tokenization for Efficient LLMs Adapted from LZW Compression

ResearchDGX agent

arXiv:2410.21548v3 Announce Type: replace Abstract: Large language models have drastically changed the prospects of AI by introducing technologies for more complex natural language processing. However

Neural Recovery of Historical Lexical Structure in Bantu Languages from Modern Data

ResearchDGX agent

arXiv:2604.22730v1 Announce Type: cross Abstract: We investigate whether neural models trained exclusively on modern morphological data can recover cross-lingual lexical structure consistent with hist

Non-Minimal Sampling and Consensus for Prohibitively Large Datasets

ResearchDGX agent

arXiv:2604.22518v1 Announce Type: new Abstract: We introduce NONSAC (Non-Minimal Sampling and Consensus), a general framework for robust and scalable model estimation from arbitrarily large datasets c

Parameter-Efficient Conditioning for Material Generalization in Graph-Based Simulators

Model ReleasesDGX agent

arXiv:2511.05456v2 Announce Type: replace Abstract: Graph network-based simulators (GNS) have demonstrated strong potential for learning particle-based physics (such as fluids, deformable solids, and

PSI: A Benchmark for Human Interpretation and Response in Traffic Interactions

Model ReleasesDGX agent

arXiv:2112.02604v3 Announce Type: replace-cross Abstract: Accurately modeling pedestrian intention and understanding driver decision-making processes are critical for the development of safe and socia

Regularized Meta-Learning for Improved Generalization

Model ReleasesDGX agent

arXiv:2602.12469v2 Announce Type: replace Abstract: Deep ensemble methods often improve predictive performance, yet they suffer from three practical limitations: redundancy among base models that infl

Removing Sandbagging in LLMs by Training with Weak Supervision

ResearchDGX agent

arXiv:2604.22082v1 Announce Type: cross Abstract: As AI systems begin to automate complex tasks, supervision increasingly relies on weaker models or limited human oversight that cannot fully verify ou

Towards Temporal Compositional Reasoning in Long-Form Sports Videos

Model ReleasesDGX agent

arXiv:2604.22226v1 Announce Type: new Abstract: Sports videos are a challenging domain for multimodal understanding because they involve complex and dynamic human activities. Despite rapid progress in

Video Analysis and Generation via a Semantic Progress Function

ApplicationsDGX agent

arXiv:2604.22554v1 Announce Type: new Abstract: Transformations produced by image and video generation models often evolve in a highly non-linear manner: long stretches where the content barely change

26 Apr 2026

I actually switched my personal Claude subscription to this (currently using Mimo v2)

Model ReleasesDGX agent

I actually switched my personal Claude subscription to this (currently using Mimo v2) Nous Portal offers everything you need to build with Hermes Agent in one easy subscription: → 300+ models from eve

THIS GUY LOST $200 IN ONE DAY BECAUSE THE STRING 'HERMES.md' WAS IN HIS GIT COMMITS HERMES.md is a real convention used in AI agent projects…

Model ReleasesDGX agent

THIS GUY LOST 200 IN ONE DAY BECAUSE THE STRING 'HERMES.md' WAS IN HIS GIT COMMITS HERMES.md is a real convention used in AI agent projects. it's a system prompt specification file. not some obscure e

24 Apr 2026

2L-LSH: A Locality-Sensitive Hash Function-Based Method For Rapid Point Cloud Indexing

ResearchDGX agent

arXiv:2604.21442v1 Announce Type: new Abstract: The development of 3D scanning technology has enabled the acquisition of massive point cloud models with diverse structures and large scales, thereby pr

A Hybridizable Neural Time Integrator for Stable Autoregressive Forecasting

ResearchDGX agent

arXiv:2604.21101v1 Announce Type: new Abstract: For autoregressive modeling of chaotic dynamical systems over long time horizons, the stability of both training and inference is a major challenge in b

AUDITA: A New Dataset to Audit Humans vs. AI Skill at Audio QA

Model ReleasesDGX agent

arXiv:2604.21766v1 Announce Type: new Abstract: Existing audio question answering benchmarks largely emphasize sound event classification or caption-grounded queries, often enabling models to succeed

Beyond Accuracy: A Stability-Aware Metric for Multi-Horizon Forecasting

Model ReleasesDGX agent

arXiv:2601.10863v3 Announce Type: replace Abstract: Traditional time series forecasting methods optimize for accuracy alone. This objective neglects temporal consistency, in other words, how consisten

Can MLLMs 'Read' What is Missing?

Model ReleasesDGX agent

arXiv:2604.21277v1 Announce Type: new Abstract: We introduce MMTR-Bench, a benchmark designed to evaluate the intrinsic ability of Multimodal Large Language Models (MLLMs) to reconstruct masked text d

Certified Coil Geometry Learning for Short-Range Magnetic Actuation and Spacecraft Docking Application

ResearchDGX agent

arXiv:2507.03806v3 Announce Type: replace-cross Abstract: This paper presents a learning-based framework for approximating an exact magnetic-field interaction model, supported by both numerical and ex

Continuous-Utility Direct Preference Optimization

SafetyDGX agent

arXiv:2602.00931v2 Announce Type: replace-cross Abstract: Large language model reasoning is often treated as a monolithic capability, relying on binary preference supervision that fails to capture par

DAVIS: OOD Detection via Dominant Activations and Variance for Increased Separation

Model ReleasesDGX agent

arXiv:2601.22703v2 Announce Type: replace Abstract: Detecting out-of-distribution (OOD) inputs is a critical safeguard for deploying machine learning models in the real world. However, most post-hoc d

Empirical Comparison of Agent Communication Protocols for Task Orchestration

Model ReleasesDGX agent

arXiv:2603.22823v3 Announce Type: replace Abstract: Context. The problem of comparative evaluation of communication protocols for task orchestration by large language model (LLM) agents is considered.

EVENT5Ws: A Large Dataset for Open-Domain Event Extraction from Documents

Model ReleasesDGX agent

arXiv:2604.21890v1 Announce Type: new Abstract: Event extraction identifies the central aspects of events from text. It supports event understanding and analysis, which is crucial for tasks such as in

Finding Meaning in Embeddings: Concept Separation Curves

ResearchDGX agent

arXiv:2604.21555v1 Announce Type: new Abstract: Sentence embedding techniques aim to encode key concepts of a sentence's meaning in a vector space. However, the majority of evaluation approaches for s

Frequency-Forcing: From Scaling-as-Time to Soft Frequency Guidance

ResearchDGX agent

arXiv:2604.20902v1 Announce Type: cross Abstract: While standard flow-matching models transport noise to data uniformly, incorporating an explicit generation order - specifically, establishing coarse,

From Codebooks to VLMs: Evaluating Automated Visual Discourse Analysis for Climate Change on Social Media

Model ReleasesDGX agent

arXiv:2604.21786v1 Announce Type: new Abstract: Social media platforms have become primary arenas for climate communication, generating millions of images and posts that - if systematically analysed -

Generalizing Numerical Reasoning in Table Data through Operation Sketches and Self-Supervised Learning

Model ReleasesDGX agent

arXiv:2604.21495v1 Announce Type: cross Abstract: Numerical reasoning over expert-domain tables often exhibits high in-domain accuracy but limited robustness to domain shift. Models trained with super

Geometric Characterisation and Structured Trajectory Surrogates for Clinical Dataset Condensation

Model ReleasesDGX agent

arXiv:2604.21638v1 Announce Type: new Abstract: Dataset condensation constructs compact synthetic datasets that retain the training utility of large real-world datasets, enabling efficient model devel

GPT-5.5 is now available in Devin as an Agent Preview! GPT-5.5 has set a new bar for what's possible with Devin. It runs longer and more aut…

Model ReleasesDGX agent

GPT-5.5 is now available in Devin as an Agent Preview! GPT-5.5 has set a new bar for what's possible with Devin. It runs longer and more autonomously than any GPT model we've tested, surfacing bugs no

Grok Voice is used by @Starlink

ApplicationsDGX agent

Grok Voice is used by @Starlink Introducing Grok Voice Think Fast 1.0 A state-of-the-art voice model built for complex, multi-step workflows with snappy responses and high accuracy. It takes the top s

← Previous
1…318319320321322…1041
Next →