AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
All
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,553 results
Model Releases

Parameterized Quantum Circuits as Feature Maps: Representation Quality and Readout Effects in Multispectral Land-Cover Classification

DGX agent

arXiv:2604.26675v1 Announce Type: cross Abstract: We investigate variational quantum classifiers (VQCs) for land-cover classification from multispectral satellite imagery, adopting a feature-map persp

model-releasesarxiv-cs-lg
30 Apr 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

PATCH: Learnable Tile-level Hybrid Sparsity for LLMs

DGX agent

arXiv:2509.23410v4 Announce Type: replace-cross Abstract: Large language models (LLMs) deliver impressive performance but incur prohibitive memory and compute costs at deployment. Model pruning is an

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

PEOPLE ARE NOW RUNNING CLAUDE CODE WITH LOCAL AI MODELS TO AVOID API COSTS. By connecting tools like Ollama and Gemma 4, developers can buil…

DGX agent

PEOPLE ARE NOW RUNNING CLAUDE CODE WITH LOCAL AI MODELS TO AVOID API COSTS. By connecting tools like Ollama and Gemma 4, developers can build apps locally with unlimited usage and no monthly billing.

model-releasesollama--x
30 Apr 2026
Model Releases

Perception Test 2025: Challenge Summary and a Unified VQA Extension

DGX agent

arXiv:2601.06287v2 Announce Type: replace Abstract: The Third Perception Test challenge was organised as a full-day workshop alongside the IEEE/CVF International Conference on Computer Vision (ICCV) 2

model-releasesarxiv-cs-cv
30 Apr 2026
Model Releases

Preserving Disagreement: Architectural Heterogeneity and Coherence Validation in Multi-Agent Policy Simulation

DGX agent

arXiv:2604.26561v1 Announce Type: cross Abstract: Multi-agent deliberation systems using large language models (LLMs) are increasingly proposed for policy simulation, yet they suffer from artificial c

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

Progressive Semantic Communication for Efficient Edge-Cloud Vision-Language Models

DGX agent

arXiv:2604.26508v1 Announce Type: cross Abstract: Deploying Vision-Language Models (VLMs) on edge devices remains challenging due to their substantial computational and memory demands, which exceed th

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

QERNEL: a Scalable Large Electron Model

DGX agent

arXiv:2604.26018v1 Announce Type: cross Abstract: We introduce QERNEL, a foundational neural wavefunction that variationally solves families of parameterized many-electron Hamiltonians and captures th

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

Quantum Feature Selection with Higher-Order Binary Optimization on Trapped-Ion Hardware

DGX agent

arXiv:2604.26834v1 Announce Type: cross Abstract: We present a quantum feature-selection framework based on a higher-order unconstrained binary optimization (HUBO) formulation that explicitly incorpor

model-releasesarxiv-cs-lg
30 Apr 2026
Model Releases

QYOLO: Lightweight Object Detection via Quantum Inspired Shared Channel Mixing

DGX agent

arXiv:2604.26435v1 Announce Type: cross Abstract: The rapid advancement of object detection architectures has positioned single stage detectors as the dominant solution for real-time visual perception

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

RADIO-ViPE: Online Tightly Coupled Multi-Modal Fusion for Open-Vocabulary Semantic SLAM in Dynamic Environments

DGX agent

arXiv:2604.26067v1 Announce Type: new Abstract: We present RADIO-ViPE (Reduce All Domains Into One -- Video Pose Engine), an online semantic SLAM system that enables geometry-aware open-vocabulary gro

model-releasesarxiv-cs-cv
30 Apr 2026
Model Releases

RaMP: Runtime-Aware Megakernel Polymorphism for Mixture-of-Experts

DGX agent

arXiv:2604.26039v1 Announce Type: cross Abstract: The optimal kernel configuration for Mixture-of-Experts (MoE) inference depends on both batch size and the expert routing distribution, yet production

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

Random Cloud: Finding Minimal Neural Architectures Without Training

DGX agent

arXiv:2604.26830v1 Announce Type: cross Abstract: I propose the Random Cloud method, a training-free approach to neural architecture search that discovers minimal feedforward network topologies throug

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

Reasoning Gets Harder for LLMs Inside A Dialogue

DGX agent

arXiv:2603.20133v2 Announce Type: replace Abstract: Large Language Models (LLMs) achieve strong performance on many reasoning benchmarks, yet these evaluations typically focus on isolated tasks that d

model-releasesarxiv-cs-cl
30 Apr 2026
Model Releases

ReLoop: Structured Modeling and Behavioral Verification for Reliable LLM-Based Optimization

DGX agent

arXiv:2602.15983v2 Announce Type: replace-cross Abstract: Large language models (LLMs) can translate natural language into optimization code, but silent failures pose a critical risk: code that execut

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

resharing this note, find it helpful given all the great open evals work + teams building vertical agents Evals are a proxy for the behavior…

DGX agent

resharing this note, find it helpful given all the great open evals work + teams building vertical agents Evals are a proxy for the behavior we want our agent to exhibit in production Model+Harness pu

model-releasesharrison-chase--x
30 Apr 2026
Model Releases

Retrieval-Augmented LLMs for Evidence Localization in Clinical Trial Recruitment from Longitudinal EHR Narratives

DGX agent

arXiv:2604.05190v2 Announce Type: replace-cross Abstract: Screening patients for enrollment is a well-known, labor-intensive bottleneck that leads to under-enrollment and, ultimately, trial failures.

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

reward-lens: A Mechanistic Interpretability Library for Reward Models

DGX agent

arXiv:2604.26130v1 Announce Type: cross Abstract: Every RLHF-trained language model is shaped by a reward model, yet the mechanistic interpretability toolkit -- logit lens, direct logit attribution, a

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

Runpod launches Flash to bring AI inference to developers without infra overhead

DGX agent

Developer-centered artificial intelligence cloud provider Runpod Inc. today announced the launch of Flash, a software development kit and platform that removes the infrastructure overhead for deployin

model-releasessiliconangle
30 Apr 2026
Model Releases

Safety Is Not Universal: The Selective Safety Trap in LLM Alignment

DGX agent

arXiv:2601.04389v2 Announce Type: replace-cross Abstract: Current safety evaluations of large language models (LLMs) create a dangerous illusion of universal protection by aggregating harms under gene

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

Say hello to the Cohere Centre! 🇨🇦 We’re proud to partner with Ottawa’s premier convention and event facility as it enters an exciting new…

DGX agent

Say hello to the Cohere Centre! 🇨🇦 We’re proud to partner with Ottawa’s premier convention and event facility as it enters an exciting new chapter. The Cohere Centre will serve as a hub where leaders

model-releasescohere--x
30 Apr 2026
Model Releases

SciMDR: Advancing Scientific Multimodal Document Reasoning

DGX agent

arXiv:2603.12249v2 Announce Type: replace-cross Abstract: Constructing scientific multimodal document reasoning datasets for foundation model training involves an inherent trade-off among scale, faith

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

SEAL: Semantic-aware Single-image Sticker Personalization with a Large-scale Sticker-tag Dataset

DGX agent

arXiv:2604.26883v1 Announce Type: new Abstract: Synthesizing a target concept from a single reference image is challenging in diffusion-based personalized text-to-image generation, particularly for st

model-releasesarxiv-cs-cv
30 Apr 2026
Model Releases

Self-Jailbreaking: Language Models Can Reason Themselves Out of Safety Alignment After Benign Reasoning Training

DGX agent

arXiv:2510.20956v2 Announce Type: replace-cross Abstract: We discover a novel and surprising phenomenon of unintentional misalignment in reasoning language models (RLMs), which we call self-jailbreaki

model-releasesarxiv-cs-cl
30 Apr 2026
Model Releases

Shorthand for Thought: Compressing LLM Reasoning via Entropy-Guided Supertokens

DGX agent

arXiv:2604.26355v1 Announce Type: new Abstract: Reasoning in Large Language Models incurs significant inference-time compute, yet the token-level information structure of reasoning traces remains unde

model-releasesarxiv-cs-cl
30 Apr 2026
Model Releases

Single unedited raw photo vs processed stack, captured from behind the moon. The color is there, just so faint it takes dozens of photos to …

DGX agent

Single unedited raw photo vs processed stack, captured from behind the moon. The color is there, just so faint it takes dozens of photos to extract Lunar photography to the absolute extreme Just a hin

model-releasesanthropic--x
30 Apr 2026
Model Releases

SongBench: A Fine-Grained Multi-Aspect Benchmark for Song Quality Assessment

DGX agent

arXiv:2604.25937v1 Announce Type: cross Abstract: Recent advancements in Text-to-Song generation have enabled realistic musical content production, yet existing evaluation benchmarks lack the professi

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

StarDrinks: An English and Korean Test Set for SLU Evaluation in a Drink Ordering Scenario

DGX agent

arXiv:2604.26500v1 Announce Type: new Abstract: LLMs and speech assistants are increasingly used for task-oriented interactions, yet their evaluation often relies on controlled scenarios that fail to

model-releasesarxiv-cs-cl
30 Apr 2026
Model Releases

State Beyond Appearance: Diagnosing and Improving State Consistency in Dial-Based Measurement Reading

DGX agent

arXiv:2604.26614v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have achieved impressive progress on general multimodal tasks, yet they remain brittle on dial-based measuremen

model-releasesarxiv-cs-cv
30 Apr 2026
Model Releases

StratMem-Bench: Evaluating Strategic Memory Use in Virtual Character Conversation Beyond Factual Recall

DGX agent

arXiv:2604.26243v1 Announce Type: cross Abstract: Achieving realistic human-like conversation for virtual characters requires not only a simple memorization and recall of past events, but also the str

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

Stress Testing Factual Consistency Metrics for Long-Document Summarization

DGX agent

arXiv:2511.07689v2 Announce Type: replace-cross Abstract: Evaluating the factual consistency of abstractive text summarization remains a significant challenge, particularly for long documents, where c

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

Structural Generalization on SLOG without Hand-Written Rules

DGX agent

arXiv:2604.26157v1 Announce Type: cross Abstract: Structural generalization in semantic parsing requires systems to apply learned compositional rules to novel structural combinations. Existing approac

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

SWE-Edit: Rethinking Code Editing for Efficient SWE-Agent

DGX agent

arXiv:2604.26102v1 Announce Type: cross Abstract: Large language model agents have achieved remarkable progress on software engineering tasks, yet current approaches suffer from a fundamental context

model-releasesarxiv-cs-cl
30 Apr 2026
Model Releases

TAP into the Patch Tokens: Leveraging Vision Foundation Model Features for AI-Generated Image Detection

DGX agent

arXiv:2604.26772v1 Announce Type: new Abstract: Recent methods demonstrate that large-scale pretrained models, such as CLIP vision transformers, effectively detect AI-generated images (AIGIs) from uns

model-releasesarxiv-cs-cv
30 Apr 2026
Model Releases

The Hermes Agent Creative Hackathon ends on Sunday, just enough time for a last minute weekend project! Here's a thread of some of the power…

DGX agent

The Hermes Agent Creative Hackathon ends on Sunday, just enough time for a last minute weekend project! Here's a thread of some of the powerful creative skills and tools we've released that might help

model-releasesnous-research--x
30 Apr 2026
Model Releases

The model wars are over. Now, Google is fighting for something bigger

DGX agent

Whoever controls the agentic control plane controls enterprise AI. Google LLC just showed up to that fight with everything it has. The company claiming it can own the full stack arrived at Google Clou

model-releasessiliconangle
30 Apr 2026
Model Releases

The Prompt Engineering Report Distilled: Quick Start Guide for Life Sciences

DGX agent

arXiv:2509.11295v2 Announce Type: replace Abstract: Developing effective prompts demands significant cognitive investment to generate reliable, high-quality responses from Large Language Models (LLMs)

model-releasesarxiv-cs-cl
30 Apr 2026
Model Releases

The Unseen Adversaries: Robust and Generalized Defense Against Adversarial Patches

DGX agent

arXiv:2604.26317v1 Announce Type: new Abstract: The vulnerabilities of deep neural networks against singularities have raised serious concerns regarding their deployment in the physical world. One of

model-releasesarxiv-cs-cv
30 Apr 2026
Model Releases

Theory-Grounded Evaluation Exposes the Authorship Gap in LLM Personalization

DGX agent

arXiv:2604.26460v1 Announce Type: new Abstract: Stylistic personalization - making LLMs write in a specific individual's style, rather than merely adapting to task preferences - lacks evaluation groun

model-releasesarxiv-cs-cl
30 Apr 2026
Model Releases

Thinking with Drafting: Optical Decompression via Logical Reconstruction

DGX agent

arXiv:2602.11731v2 Announce Type: replace Abstract: Existing multimodal large language models have achieved high-fidelity visual perception and exploratory visual generation. However, a precision para

model-releasesarxiv-cs-cl
30 Apr 2026
Model Releases

This is really well thought out. Filesystems are the new default abstraction for agents to interact with documents (the new RAG stack in 202…

DGX agent

This is really well thought out. Filesystems are the new default abstraction for agents to interact with documents (the new RAG stack in 2026). The issue is actually figuring out how to productize thi

model-releasesjerry-liu--x
30 Apr 2026
Model Releases

This startup’s new mechanistic interpretability tool lets you debug LLMs

DGX agent

The San Francisco–based startup Goodfire just released a new tool, called Silico, that lets researchers and engineers peer inside an AI model and adjust its parameters—the settings that determine a mo

model-releasesmit-tech-review
30 Apr 2026
Model Releases

TildeOpen LLM: Leveraging Curriculum Learning to Achieve Equitable Language Representation

DGX agent

arXiv:2603.08182v2 Announce Type: replace-cross Abstract: Large language models often underperform in many European languages due to the dominance of English and a few high-resource languages in train

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

Time Blindness: Why Video-Language Models Can't See What Humans Can?

DGX agent

arXiv:2505.24867v2 Announce Type: replace-cross Abstract: Recent advances in vision-language models (VLMs) have made impressive strides in understanding spatio-temporal relationships in videos. Howeve

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

Time series classification with random convolution kernels: pooling operators and input representations matter

DGX agent

arXiv:2409.01115v5 Announce Type: replace Abstract: This article presents a new approach based on MiniRocket, called SelF-Rocket, for fast time series classification (TSC). Unlike existing approaches

model-releasesarxiv-cs-lg
30 Apr 2026
Model Releases

TinyR1-32B-Preview: Boosting Accuracy with Branch-Merge Distillation

DGX agent

arXiv:2503.04872v3 Announce Type: replace-cross Abstract: The challenge of reducing the size of Large Language Models (LLMs) while maintaining their performance has gained significant attention. Howev

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

Today we’re releasing Qwen-Scope 🔭, an open suite of sparse autoencoders for the Qwen model family. It turns SAE features into practical to…

DGX agent

Today we’re releasing Qwen-Scope 🔭, an open suite of sparse autoencoders for the Qwen model family. It turns SAE features into practical tools: 🎯 Inference — Steer model outputs by directly manipulati

model-releasesqwen--x
30 Apr 2026
Model Releases

Train a TensorFlow object detection model – then deploy it on a robot 🤖 Iulia Feroli (@iuliaferoli) shows how to turn a notebook into a rea…

DGX agent

Train a TensorFlow object detection model – then deploy it on a robot 🤖 Iulia Feroli (@iuliaferoli) shows how to turn a notebook into a real-time object detection app. This tutorial works for any proj

model-releasesclem-delangue--x
30 Apr 2026
Model Releases

Training-Free Adaptation of New-Generation LLMs using Legacy Clinical Models

DGX agent

arXiv:2601.03423v3 Announce Type: replace-cross Abstract: Adapting language models to the clinical domain through continued pretraining and instruction tuning requires costly retraining for each new m

model-releasesarxiv-cs-ai
30 Apr 2026
← Previous
1…374375376377378…470
Next →