AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,292 results
Model Releases

NASA wasn’t “closed” before Trump. The US doesn’t pay NATO anywhere close to “trillions of dollars” or “hundreds of billions of dollars a ye…

DGX agent

NASA wasn’t “closed” before Trump. The US doesn’t pay NATO anywhere close to “trillions of dollars” or “hundreds of billions of dollars a year.” Trump hasn’t built anywhere close to “over 1,000 miles

model-releasesanthropic--x
15 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Nemotron 3 Super: Open, Efficient Mixture-of-Experts Hybrid Mamba-Transformer Model for Agentic Reasoning

DGX agent

arXiv:2604.12374v1 Announce Type: cross Abstract: We describe the pre-training, post-training, and quantization of Nemotron 3 Super, a 120 billion (active 12 billion) parameter hybrid Mamba-Attention

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

NeuroPareto: Calibrated Acquisition for Costly Many-Goal Search in Vast Parameter Spaces

DGX agent

arXiv:2602.03901v4 Announce Type: replace Abstract: The pursuit of optimal trade-offs in high-dimensional search spaces under stringent computational constraints poses a fundamental challenge for cont

model-releasesarxiv-cs-lg
15 Apr 2026
Model Releases

NTIRE 2026 The 3rd Restore Any Image Model (RAIM) Challenge: Professional Image Quality Assessment (Track 1)

DGX agent

arXiv:2604.12512v1 Announce Type: cross Abstract: In this paper, we present an overview of the NTIRE 2026 challenge on the 3rd Restore Any Image Model in the Wild, specifically focusing on Track 1: Pr

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Nucleus-Image: Sparse MoE for Image Generation

DGX agent

arXiv:2604.12163v1 Announce Type: new Abstract: We present Nucleus-Image, a text-to-image generation model that establishes a new Pareto frontier in quality-versus-efficiency by matching or exceeding

model-releasesarxiv-cs-cv
15 Apr 2026
Model Releases

OFA-Diffusion Compression: Compressing Diffusion Model in One-Shot Manner

DGX agent

arXiv:2604.12668v1 Announce Type: new Abstract: The Diffusion Probabilistic Model (DPM) achieves remarkable performance in image generation, while its increasing parameter size and computational overh

model-releasesarxiv-cs-cv
15 Apr 2026
Model Releases

ollama launch claude --model glm-5.1:cloud (We are rushing to get more capacity 🙏🙏🙏)

DGX agent

ollama launch claude --model glm-5.1:cloud (We are rushing to get more capacity 🙏🙏🙏) Uber's CTO told @LauraBratton5 that AI coding tools—particularly Anthropic’s Claude Code—has already maxed out its

model-releasesollama--x
15 Apr 2026
Model Releases

Olmo 3

DGX agent

arXiv:2512.13961v2 Announce Type: replace Abstract: We introduce Olmo 3, a family of state-of-the-art, fully-open language models at the 7B and 32B parameter scales. Olmo 3 model construction targets

model-releasesarxiv-cs-cl
15 Apr 2026
Model Releases

OmniHands: Towards Robust 4D Hand Mesh Recovery via A Versatile Transformer

DGX agent

arXiv:2405.20330v4 Announce Type: replace-cross Abstract: In this paper, we introduce OmniHands, a universal approach to recovering interactive hand meshes and their relative movement from monocular o

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

On Higher-Order Geometric Refinements of Classical Covariance Asymptotics: An Approach via Intrinsic and Extrinsic Information Geometry

DGX agent

arXiv:2604.12725v1 Announce Type: cross Abstract: Classical Fisher-information asymptotics describe the covariance of regular efficient estimators through the local quadratic approximation of the log-

model-releasesarxiv-cs-lg
15 Apr 2026
Model Releases

On the continuum limit of t-SNE for data visualization

DGX agent

arXiv:2604.12041v1 Announce Type: cross Abstract: This work is concerned with the continuum limit of a graph-based data visualization technique called the t-Distributed Stochastic Neighbor Embedding (

model-releasesarxiv-cs-lg
15 Apr 2026
Model Releases

Operationalising the Right to be Forgotten in LLMs: A Lightweight Sequential Unlearning Framework for Privacy-Aligned Deployment in Politically Sensitive Environments

DGX agent

arXiv:2604.12459v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly deployed in politically sensitive environments, where memorisation of personal data or confidential conten

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Orthogonal Subspace Projection for Continual Machine Unlearning via SVD-Based LoRA

DGX agent

arXiv:2604.12526v1 Announce Type: cross Abstract: Continual machine unlearning aims to remove the influence of data that should no longer be retained, while preserving the usefulness of the model on e

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Parametric Interpolation of Dynamic Mode Decomposition for Predicting Nonlinear Systems

DGX agent

arXiv:2604.12103v1 Announce Type: cross Abstract: We present parameter-interpolated dynamic mode decomposition (piDMD), a parametric reduced-order modeling framework that embeds known parameter-affine

model-releasesarxiv-cs-lg
15 Apr 2026
Model Releases

Parcae: Scaling Laws For Stable Looped Language Models

DGX agent

arXiv:2604.12946v1 Announce Type: new Abstract: Traditional fixed-depth architectures scale quality by increasing training FLOPs, typically through increased parameterization, at the expense of a high

model-releasesarxiv-cs-lg
15 Apr 2026
Model Releases

ParetoBandit: Budget-Paced Adaptive Routing for Non-Stationary LLM Serving

DGX agent

arXiv:2604.00136v2 Announce Type: replace-cross Abstract: Multi-model LLM serving operates in a non-stationary, noisy environment: providers revise pricing, model quality can shift or regress without

model-releasesarxiv-cs-cl
15 Apr 2026
Model Releases

Parsing complex tables in PDFs is extremely challenging. Existing metrics for measuring table accuracy, like TEDS (tree edit distance simila…

DGX agent

Parsing complex tables in PDFs is extremely challenging. Existing metrics for measuring table accuracy, like TEDS (tree edit distance similarity), overweight exact table structure and underweight sema

model-releasesjerry-liu--x
15 Apr 2026
Model Releases

Physically Accurate Rigid-Body Dynamics in Particle-Based Simulation

DGX agent

arXiv:2603.14634v3 Announce Type: replace Abstract: Robotics demands simulation that can reason about the diversity of real-world physical interactions, from rigid to deformable objects and fluids. Cu

model-releasesarxiv-cs-ro
15 Apr 2026
Model Releases

Policy-Invisible Violations in LLM-Based Agents

DGX agent

arXiv:2604.12177v1 Announce Type: new Abstract: LLM-based agents can execute actions that are syntactically valid, user-sanctioned, and semantically appropriate, yet still violate organizational polic

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

PolicyLLM: Towards Excellent Comprehension of Public Policy for Large Language Models

DGX agent

arXiv:2604.12995v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly integrated into real-world decision-making, including in the domain of public policy. Yet, their ability t

model-releasesarxiv-cs-cl
15 Apr 2026
Model Releases

Polynomial Expansion Rank Adaptation: Enhancing Low-Rank Fine-Tuning with High-Order Interactions

DGX agent

arXiv:2604.11841v1 Announce Type: cross Abstract: Low-rank adaptation (LoRA) is a widely used strategy for efficient fine-tuning of large language models (LLMs), but its strictly linear structure fund

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

ProbeLogits: Kernel-Level LLM Inference Primitives for AI-Native Operating Systems

DGX agent

arXiv:2604.11943v1 Announce Type: cross Abstract: An OS kernel that runs LLM inference internally can read logit distributions before any text is generated -- and act on them as a governance primitive

model-releasesarxiv-cs-lg
15 Apr 2026
Model Releases

PromptEcho: Annotation-Free Reward from Vision-Language Models for Text-to-Image Reinforcement Learning

DGX agent

arXiv:2604.12652v1 Announce Type: cross Abstract: Reinforcement learning (RL) can improve the prompt following capability of text-to-image (T2I) models, yet obtaining high-quality reward signals remai

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Public Profile Matters: A Scalable Integrated Approach to Recommend Citations in the Wild

DGX agent

arXiv:2603.17361v2 Announce Type: replace-cross Abstract: Proper citation of relevant literature is essential for contextualising and validating scientific contributions. While current citation recomm

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Quantile Q-Learning: Revisiting Offline Extreme Q-Learning with Quantile Regression

DGX agent

arXiv:2511.11973v2 Announce Type: replace Abstract: Offline reinforcement learning (RL) enables policy learning from fixed datasets without further environment interaction, making it particularly valu

model-releasesarxiv-cs-lg
15 Apr 2026
Model Releases

QuarkMedSearch: A Long-Horizon Deep Search Agent for Exploring Medical Intelligence

DGX agent

arXiv:2604.12867v1 Announce Type: new Abstract: As agentic foundation models continue to evolve, how to further improve their performance in vertical domains has become an important challenge. To this

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Ran ChatGPT Plus and Claude Pro side by side for 30 days, here's what I found as a daily ChatGPT user

DGX agent

A Reddit post from r/ChatGPT in which a longtime ChatGPT Plus user shares findings after running both ChatGPT Plus and Claude Pro simultaneously for 30 days. The post likely reflects a real-world comp

model-releasesr-chatgpt
15 Apr 2026
Model Releases

RankOOD -- Class Ranking-based Out-of-Distribution Detection

DGX agent

arXiv:2511.19996v2 Announce Type: replace Abstract: We propose RankOOD, a rank-based Out-of-Distribution (OOD) detection approach based on training a model with the Placket-Luce loss, which is now ext

model-releasesarxiv-cs-lg
15 Apr 2026
Model Releases

Reading Between the Pixels: Linking Text-Image Embedding Alignment to Typographic Attack Success on Vision-Language Models

DGX agent

arXiv:2604.12371v1 Announce Type: new Abstract: We study typographic prompt injection attacks on vision-language models (VLMs), where adversarial text is rendered as images to bypass safety mechanisms

model-releasesarxiv-cs-cv
15 Apr 2026
Model Releases

ReasonXL: Shifting LLM Reasoning Language Without Sacrificing Performance

DGX agent

arXiv:2604.12378v1 Announce Type: new Abstract: Despite advances in multilingual capabilities, most large language models (LLMs) remain English-centric in their training and, crucially, in their produ

model-releasesarxiv-cs-cl
15 Apr 2026
Model Releases

Red Teaming Large Reasoning Models

DGX agent

arXiv:2512.00412v4 Announce Type: replace-cross Abstract: Large Reasoning Models (LRMs) have emerged as a powerful advancement in multi-step reasoning tasks, offering enhanced transparency and logical

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

ReflectCAP: Detailed Image Captioning with Reflective Memory

DGX agent

arXiv:2604.12357v1 Announce Type: new Abstract: Detailed image captioning demands both factual grounding and fine-grained coverage, yet existing methods have struggled to achieve them simultaneously.

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Revisiting the Reliability of Language Models in Instruction-Following

DGX agent

arXiv:2512.14754v2 Announce Type: replace-cross Abstract: Advanced LLMs have achieved near-ceiling instruction-following accuracy on benchmarks such as IFEval. However, these impressive scores do not

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Robust Explanations for User Trust in Enterprise NLP Systems

DGX agent

arXiv:2604.12069v1 Announce Type: cross Abstract: Robust explanations are increasingly required for user trust in enterprise NLP, yet pre-deployment validation is difficult in the common case of black

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Round-Trip Translation Reveals What Frontier Multilingual Benchmarks Miss

DGX agent

arXiv:2604.12911v1 Announce Type: cross Abstract: Multilingual benchmarks guide the development of frontier models. Yet multilingual evaluations reported by frontier models are structured similar to p

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

RPG-SAM: Reliability-Weighted Prototypes and Geometric Adaptive Threshold Selection for Training-Free One-Shot Polyp Segmentation

DGX agent

arXiv:2603.07436v2 Announce Type: replace Abstract: Training-free one-shot segmentation offers a scalable alternative to expert annotations where knowledge is often transferred from support images and

model-releasesarxiv-cs-cv
15 Apr 2026
Model Releases

Safe-SAIL: Towards a Fine-grained Safety Landscape of Large Language Models via Sparse Autoencoder Interpretation Framework

DGX agent

arXiv:2509.18127v3 Announce Type: replace-cross Abstract: Sparse autoencoders (SAEs) enable interpretability research by decomposing entangled model activations into monosemantic features. However, un

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Scale-aware Message Passing For Graph Node Classification

DGX agent

arXiv:2411.19392v3 Announce Type: replace Abstract: Most Graph Neural Networks (GNNs) operate at the first-order scale, even though multi-scale representations are known to be crucial in domains such

model-releasesarxiv-cs-lg
15 Apr 2026
Model Releases

SEATrack: Simple, Efficient, and Adaptive Multimodal Tracker

DGX agent

arXiv:2604.12502v1 Announce Type: cross Abstract: Parameter-efficient fine-tuning (PEFT) in multimodal tracking reveals a concerning trend where recent performance gains are often achieved at the cost

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

SecureWebArena: A Holistic Security Evaluation Benchmark for LVLM-based Web Agents

DGX agent

arXiv:2510.10073v2 Announce Type: replace-cross Abstract: Large vision-language model (LVLM)-based web agents are emerging as powerful tools for automating complex online tasks. However, when deployed

model-releasesarxiv-cs-cv
15 Apr 2026
Model Releases

See, Point, Refine: Multi-Turn Approach to GUI Grounding with Visual Feedback

DGX agent

arXiv:2604.13019v1 Announce Type: new Abstract: Computer Use Agents (CUAs) fundamentally rely on graphical user interface (GUI) grounding to translate language instructions into executable screen acti

model-releasesarxiv-cs-cv
15 Apr 2026
Model Releases

SeedPrints: Fingerprints Can Even Tell Which Seed Your Large Language Model Was Trained From

DGX agent

arXiv:2509.26404v2 Announce Type: replace-cross Abstract: Fingerprinting Large Language Models (LLMs)is essential for provenance verification and model attribution. Existing fingerprinting methods are

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Self-Adversarial One Step Generation via Condition Shifting

DGX agent

arXiv:2604.12322v1 Announce Type: new Abstract: The push for efficient text to image synthesis has moved the field toward one step sampling, yet existing methods still face a three way tradeoff among

model-releasesarxiv-cs-cv
15 Apr 2026
Model Releases

Self-Monitoring Benefits from Structural Integration: Lessons from Metacognition in Continuous-Time Multi-Timescale Agents

DGX agent

arXiv:2604.11914v1 Announce Type: new Abstract: Self-monitoring capabilities -- metacognition, self-prediction, and subjective duration -- are often proposed as useful additions to reinforcement learn

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

SEW: Self-Evolving Agentic Workflows for Automated Code Generation

DGX agent

arXiv:2505.18646v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have demonstrated effectiveness in code generation tasks. To enable LLMs to address more complex coding challenge

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Silo-Bench: A Scalable Environment for Evaluating Distributed Coordination in Multi-Agent LLM Systems

DGX agent

arXiv:2603.01045v2 Announce Type: replace-cross Abstract: Large language models are increasingly deployed in multi-agent systems to overcome context limitations by distributing information across agen

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

SinkSAM-Net: Knowledge-Driven Self-Supervised Sinkhole Segmentation Using Topographic Priors and Segment Anything Model

DGX agent

arXiv:2410.01473v2 Announce Type: replace Abstract: Soil sinkholes significantly influence soil degradation, infrastructure vulnerability, and landscape evolution. However, their irregular shapes, com

model-releasesarxiv-cs-cv
15 Apr 2026
Model Releases

SIR-Bench: Evaluating Investigation Depth in Security Incident Response Agents

DGX agent

arXiv:2604.12040v1 Announce Type: cross Abstract: We present SIR-Bench, a benchmark of 794 test cases for evaluating autonomous security incident response agents that distinguishes genuine forensic in

model-releasesarxiv-cs-ai
15 Apr 2026
← Previous
1…434435436437438…465
Next →