AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,545 results
Model Releases

QYOLO: Lightweight Object Detection via Quantum Inspired Shared Channel Mixing

DGX agent

arXiv:2604.26435v1 Announce Type: cross Abstract: The rapid advancement of object detection architectures has positioned single stage detectors as the dominant solution for real-time visual perception

model-releasesarxiv-cs-ai
30 Apr 2026
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

RADIO-ViPE: Online Tightly Coupled Multi-Modal Fusion for Open-Vocabulary Semantic SLAM in Dynamic Environments

DGX agent

arXiv:2604.26067v1 Announce Type: new Abstract: We present RADIO-ViPE (Reduce All Domains Into One -- Video Pose Engine), an online semantic SLAM system that enables geometry-aware open-vocabulary gro

model-releasesarxiv-cs-cv
30 Apr 2026
Model Releases

RaMP: Runtime-Aware Megakernel Polymorphism for Mixture-of-Experts

DGX agent

arXiv:2604.26039v1 Announce Type: cross Abstract: The optimal kernel configuration for Mixture-of-Experts (MoE) inference depends on both batch size and the expert routing distribution, yet production

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

Random Cloud: Finding Minimal Neural Architectures Without Training

DGX agent

arXiv:2604.26830v1 Announce Type: cross Abstract: I propose the Random Cloud method, a training-free approach to neural architecture search that discovers minimal feedforward network topologies throug

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

Reasoning Gets Harder for LLMs Inside A Dialogue

DGX agent

arXiv:2603.20133v2 Announce Type: replace Abstract: Large Language Models (LLMs) achieve strong performance on many reasoning benchmarks, yet these evaluations typically focus on isolated tasks that d

model-releasesarxiv-cs-cl
30 Apr 2026
Model Releases

ReLoop: Structured Modeling and Behavioral Verification for Reliable LLM-Based Optimization

DGX agent

arXiv:2602.15983v2 Announce Type: replace-cross Abstract: Large language models (LLMs) can translate natural language into optimization code, but silent failures pose a critical risk: code that execut

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

resharing this note, find it helpful given all the great open evals work + teams building vertical agents Evals are a proxy for the behavior…

DGX agent

resharing this note, find it helpful given all the great open evals work + teams building vertical agents Evals are a proxy for the behavior we want our agent to exhibit in production Model+Harness pu

model-releasesharrison-chase--x
30 Apr 2026
Model Releases

Retrieval-Augmented LLMs for Evidence Localization in Clinical Trial Recruitment from Longitudinal EHR Narratives

DGX agent

arXiv:2604.05190v2 Announce Type: replace-cross Abstract: Screening patients for enrollment is a well-known, labor-intensive bottleneck that leads to under-enrollment and, ultimately, trial failures.

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

reward-lens: A Mechanistic Interpretability Library for Reward Models

DGX agent

arXiv:2604.26130v1 Announce Type: cross Abstract: Every RLHF-trained language model is shaped by a reward model, yet the mechanistic interpretability toolkit -- logit lens, direct logit attribution, a

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

Runpod launches Flash to bring AI inference to developers without infra overhead

DGX agent

Developer-centered artificial intelligence cloud provider Runpod Inc. today announced the launch of Flash, a software development kit and platform that removes the infrastructure overhead for deployin

model-releasessiliconangle
30 Apr 2026
Model Releases

Safety Is Not Universal: The Selective Safety Trap in LLM Alignment

DGX agent

arXiv:2601.04389v2 Announce Type: replace-cross Abstract: Current safety evaluations of large language models (LLMs) create a dangerous illusion of universal protection by aggregating harms under gene

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

Say hello to the Cohere Centre! 🇨🇦 We’re proud to partner with Ottawa’s premier convention and event facility as it enters an exciting new…

DGX agent

Say hello to the Cohere Centre! 🇨🇦 We’re proud to partner with Ottawa’s premier convention and event facility as it enters an exciting new chapter. The Cohere Centre will serve as a hub where leaders

model-releasescohere--x
30 Apr 2026
Model Releases

SciMDR: Advancing Scientific Multimodal Document Reasoning

DGX agent

arXiv:2603.12249v2 Announce Type: replace-cross Abstract: Constructing scientific multimodal document reasoning datasets for foundation model training involves an inherent trade-off among scale, faith

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

SEAL: Semantic-aware Single-image Sticker Personalization with a Large-scale Sticker-tag Dataset

DGX agent

arXiv:2604.26883v1 Announce Type: new Abstract: Synthesizing a target concept from a single reference image is challenging in diffusion-based personalized text-to-image generation, particularly for st

model-releasesarxiv-cs-cv
30 Apr 2026
Model Releases

Self-Jailbreaking: Language Models Can Reason Themselves Out of Safety Alignment After Benign Reasoning Training

DGX agent

arXiv:2510.20956v2 Announce Type: replace-cross Abstract: We discover a novel and surprising phenomenon of unintentional misalignment in reasoning language models (RLMs), which we call self-jailbreaki

model-releasesarxiv-cs-cl
30 Apr 2026
Model Releases

Shorthand for Thought: Compressing LLM Reasoning via Entropy-Guided Supertokens

DGX agent

arXiv:2604.26355v1 Announce Type: new Abstract: Reasoning in Large Language Models incurs significant inference-time compute, yet the token-level information structure of reasoning traces remains unde

model-releasesarxiv-cs-cl
30 Apr 2026
Model Releases

Single unedited raw photo vs processed stack, captured from behind the moon. The color is there, just so faint it takes dozens of photos to …

DGX agent

Single unedited raw photo vs processed stack, captured from behind the moon. The color is there, just so faint it takes dozens of photos to extract Lunar photography to the absolute extreme Just a hin

model-releasesanthropic--x
30 Apr 2026
Model Releases

SongBench: A Fine-Grained Multi-Aspect Benchmark for Song Quality Assessment

DGX agent

arXiv:2604.25937v1 Announce Type: cross Abstract: Recent advancements in Text-to-Song generation have enabled realistic musical content production, yet existing evaluation benchmarks lack the professi

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

StarDrinks: An English and Korean Test Set for SLU Evaluation in a Drink Ordering Scenario

DGX agent

arXiv:2604.26500v1 Announce Type: new Abstract: LLMs and speech assistants are increasingly used for task-oriented interactions, yet their evaluation often relies on controlled scenarios that fail to

model-releasesarxiv-cs-cl
30 Apr 2026
Model Releases

State Beyond Appearance: Diagnosing and Improving State Consistency in Dial-Based Measurement Reading

DGX agent

arXiv:2604.26614v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have achieved impressive progress on general multimodal tasks, yet they remain brittle on dial-based measuremen

model-releasesarxiv-cs-cv
30 Apr 2026
Model Releases

StratMem-Bench: Evaluating Strategic Memory Use in Virtual Character Conversation Beyond Factual Recall

DGX agent

arXiv:2604.26243v1 Announce Type: cross Abstract: Achieving realistic human-like conversation for virtual characters requires not only a simple memorization and recall of past events, but also the str

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

Stress Testing Factual Consistency Metrics for Long-Document Summarization

DGX agent

arXiv:2511.07689v2 Announce Type: replace-cross Abstract: Evaluating the factual consistency of abstractive text summarization remains a significant challenge, particularly for long documents, where c

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

Structural Generalization on SLOG without Hand-Written Rules

DGX agent

arXiv:2604.26157v1 Announce Type: cross Abstract: Structural generalization in semantic parsing requires systems to apply learned compositional rules to novel structural combinations. Existing approac

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

SWE-Edit: Rethinking Code Editing for Efficient SWE-Agent

DGX agent

arXiv:2604.26102v1 Announce Type: cross Abstract: Large language model agents have achieved remarkable progress on software engineering tasks, yet current approaches suffer from a fundamental context

model-releasesarxiv-cs-cl
30 Apr 2026
Model Releases

TAP into the Patch Tokens: Leveraging Vision Foundation Model Features for AI-Generated Image Detection

DGX agent

arXiv:2604.26772v1 Announce Type: new Abstract: Recent methods demonstrate that large-scale pretrained models, such as CLIP vision transformers, effectively detect AI-generated images (AIGIs) from uns

model-releasesarxiv-cs-cv
30 Apr 2026
Model Releases

The Hermes Agent Creative Hackathon ends on Sunday, just enough time for a last minute weekend project! Here's a thread of some of the power…

DGX agent

The Hermes Agent Creative Hackathon ends on Sunday, just enough time for a last minute weekend project! Here's a thread of some of the powerful creative skills and tools we've released that might help

model-releasesnous-research--x
30 Apr 2026
Model Releases

The model wars are over. Now, Google is fighting for something bigger

DGX agent

Whoever controls the agentic control plane controls enterprise AI. Google LLC just showed up to that fight with everything it has. The company claiming it can own the full stack arrived at Google Clou

model-releasessiliconangle
30 Apr 2026
Model Releases

The Prompt Engineering Report Distilled: Quick Start Guide for Life Sciences

DGX agent

arXiv:2509.11295v2 Announce Type: replace Abstract: Developing effective prompts demands significant cognitive investment to generate reliable, high-quality responses from Large Language Models (LLMs)

model-releasesarxiv-cs-cl
30 Apr 2026
Model Releases

The Unseen Adversaries: Robust and Generalized Defense Against Adversarial Patches

DGX agent

arXiv:2604.26317v1 Announce Type: new Abstract: The vulnerabilities of deep neural networks against singularities have raised serious concerns regarding their deployment in the physical world. One of

model-releasesarxiv-cs-cv
30 Apr 2026
Model Releases

Theory-Grounded Evaluation Exposes the Authorship Gap in LLM Personalization

DGX agent

arXiv:2604.26460v1 Announce Type: new Abstract: Stylistic personalization - making LLMs write in a specific individual's style, rather than merely adapting to task preferences - lacks evaluation groun

model-releasesarxiv-cs-cl
30 Apr 2026
Model Releases

Thinking with Drafting: Optical Decompression via Logical Reconstruction

DGX agent

arXiv:2602.11731v2 Announce Type: replace Abstract: Existing multimodal large language models have achieved high-fidelity visual perception and exploratory visual generation. However, a precision para

model-releasesarxiv-cs-cl
30 Apr 2026
Model Releases

This is really well thought out. Filesystems are the new default abstraction for agents to interact with documents (the new RAG stack in 202…

DGX agent

This is really well thought out. Filesystems are the new default abstraction for agents to interact with documents (the new RAG stack in 2026). The issue is actually figuring out how to productize thi

model-releasesjerry-liu--x
30 Apr 2026
Model Releases

This startup’s new mechanistic interpretability tool lets you debug LLMs

DGX agent

The San Francisco–based startup Goodfire just released a new tool, called Silico, that lets researchers and engineers peer inside an AI model and adjust its parameters—the settings that determine a mo

model-releasesmit-tech-review
30 Apr 2026
Model Releases

TildeOpen LLM: Leveraging Curriculum Learning to Achieve Equitable Language Representation

DGX agent

arXiv:2603.08182v2 Announce Type: replace-cross Abstract: Large language models often underperform in many European languages due to the dominance of English and a few high-resource languages in train

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

Time Blindness: Why Video-Language Models Can't See What Humans Can?

DGX agent

arXiv:2505.24867v2 Announce Type: replace-cross Abstract: Recent advances in vision-language models (VLMs) have made impressive strides in understanding spatio-temporal relationships in videos. Howeve

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

Time series classification with random convolution kernels: pooling operators and input representations matter

DGX agent

arXiv:2409.01115v5 Announce Type: replace Abstract: This article presents a new approach based on MiniRocket, called SelF-Rocket, for fast time series classification (TSC). Unlike existing approaches

model-releasesarxiv-cs-lg
30 Apr 2026
Model Releases

TinyR1-32B-Preview: Boosting Accuracy with Branch-Merge Distillation

DGX agent

arXiv:2503.04872v3 Announce Type: replace-cross Abstract: The challenge of reducing the size of Large Language Models (LLMs) while maintaining their performance has gained significant attention. Howev

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

Today we’re releasing Qwen-Scope 🔭, an open suite of sparse autoencoders for the Qwen model family. It turns SAE features into practical to…

DGX agent

Today we’re releasing Qwen-Scope 🔭, an open suite of sparse autoencoders for the Qwen model family. It turns SAE features into practical tools: 🎯 Inference — Steer model outputs by directly manipulati

model-releasesqwen--x
30 Apr 2026
Model Releases

Train a TensorFlow object detection model – then deploy it on a robot 🤖 Iulia Feroli (@iuliaferoli) shows how to turn a notebook into a rea…

DGX agent

Train a TensorFlow object detection model – then deploy it on a robot 🤖 Iulia Feroli (@iuliaferoli) shows how to turn a notebook into a real-time object detection app. This tutorial works for any proj

model-releasesclem-delangue--x
30 Apr 2026
Model Releases

Training-Free Adaptation of New-Generation LLMs using Legacy Clinical Models

DGX agent

arXiv:2601.03423v3 Announce Type: replace-cross Abstract: Adapting language models to the clinical domain through continued pretraining and instruction tuning requires costly retraining for each new m

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

Training-Free Loosely Speculative Decoding: Accepting Semantically Correct Drafts Beyond Exact Match

DGX agent

arXiv:2511.22972v3 Announce Type: replace Abstract: Large language models (LLMs) achieve strong performance across diverse tasks but suffer from high inference latency due to their autoregressive gene

model-releasesarxiv-cs-cl
30 Apr 2026
Model Releases

tweeted about this yesterday and Cursor already dropped the alpha today! 🚀very cool to see how us, them, and others have converged on good …

DGX agent

tweeted about this yesterday and Cursor already dropped the alpha today! 🚀very cool to see how us, them, and others have converged on good design patterns in Agent + Harness Engineering: 1. Tuning dif

model-releasesharrison-chase--x
30 Apr 2026
Model Releases

Value-Guided Iterative Refinement and the DIQ-H Benchmark for Evaluating VLM Robustness

DGX agent

arXiv:2512.03992v2 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) are essential for embodied AI and safety-critical applications, such as robotics and autonomous systems. However

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

VIGNETTE: Socially Grounded Bias Evaluation for Vision-Language Models

DGX agent

arXiv:2505.22897v2 Announce Type: replace Abstract: While bias in large language models (LLMs) is well-studied, similar concerns in vision-language models (VLMs) have received comparatively less atten

model-releasesarxiv-cs-cl
30 Apr 2026
Model Releases

VLN-Cache: Enabling Token Caching for VLN Models with Visual/Semantic Dynamics Awareness

DGX agent

arXiv:2603.07080v3 Announce Type: replace-cross Abstract: Vision-and-Language Navigation (VLN) increasingly relies on large vision-language models, but their inference cost conflicts with real-time de

model-releasesarxiv-cs-lg
30 Apr 2026
Model Releases

VulStyle: A Multi-Modal Pre-Training for Code Stylometry-Augmented Vulnerability Detection

DGX agent

arXiv:2604.26313v1 Announce Type: cross Abstract: We present VulStyle, a multi-modal software vulnerability detection model that jointly encodes function-level source code, non-terminal Abstract Synta

model-releasesarxiv-cs-lg
30 Apr 2026
Model Releases

We need RSS for sharing abundant vibe-coded apps

DGX agent

We need RSS for sharing abundant vibe-coded apps Matt Webb: I would love an RSS web feed for all those various tools and apps pages, each item with an “Install” button. (But install to where?) The les

model-releasessimon-willison
30 Apr 2026
Model Releases

WebAggregator: Enhancing Compositional Reasoning Capabilities of Deep Research Agent Foundation Models

DGX agent

arXiv:2510.14438v2 Announce Type: replace Abstract: The hallmark of Deep Research agents lies in compositional reasoning, the capacity to aggregate distributed, heterogeneous information into coherent

model-releasesarxiv-cs-cl
30 Apr 2026
← Previous
1…373374375376377…470
Next →