AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
All
84,648Total entries
1Added by human
84,647Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,620 results
Model Releases

FedCritic: Serverless Federated Critic Learning-based Resource Allocation for Multi-Cell OFDMA in 6G

DGX agent

arXiv:2605.21418v1 Announce Type: cross Abstract: In sixth-generation (6G) ultra-dense networks, aggressive frequency reuse amplifies inter-cell interference (ICI), making multi-cell orthogonal freque

model-releasesarxiv-cs-cv
21 May 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Federal records: Grok was utilized in only 3 of 400+ publicly identified federal AI use cases in 2025, behind 234 for ChatGPT, 33 for Gemini, 26 for Claude (Reuters)

DGX agent

Reuters: Federal records: Grok was utilized in only 3 of 400+ publicly identified federal AI use cases in 2025, behind 234 for ChatGPT, 33 for Gemini, 26 for Claude — SpaceX's initial public offering

model-releasestechmeme
21 May 2026
Model Releases

Federated LoRA Fine-Tuning for LLMs via Collaborative Alignment

DGX agent

arXiv:2605.21217v1 Announce Type: cross Abstract: Low-rank adaptation (LoRA) has emerged as a powerful tool for parameter-efficient fine-tuning of large language models (LLMs). This paper studies LoRA

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy

DGX agent

arXiv:2605.20965v1 Announce Type: new Abstract: Large Vision-Language Models (LVLMs) have shown remarkable performance on a wide range of vision-language tasks. Despite this progress, they are still p

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

Findings of the Counter Turing Test: AI-Generated Text Detection

DGX agent

arXiv:2605.20761v1 Announce Type: new Abstract: The rapid proliferation of AI-generated text has introduced significant challenges in maintaining the integrity of digital content. Advanced generative

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

Fine-grained Claim-level RAG Benchmark for Law

DGX agent

arXiv:2605.21071v1 Announce Type: new Abstract: The rapid progress of large language models (LLMs) is shifting semantic search toward a question-answering paradigm, where users ask questions and LLMs

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

Free-Grained Hierarchical Visual Recognition

DGX agent

arXiv:2510.14737v3 Announce Type: replace Abstract: Hierarchical image recognition seeks to predict class labels along a semantic taxonomy, from broad categories to specific ones, typically under the

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

FT-Dojo: Towards Autonomous LLM Fine-Tuning with Language Agents

DGX agent

arXiv:2603.01712v2 Announce Type: replace-cross Abstract: Fine-tuning large language models for vertical domains remains labor-intensive, requiring practitioners to curate data, configure training, an

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

FullFlow: Upgrading Text-to-Image Flow Matching Models for Bidirectional Vision--Language Generation

DGX agent

arXiv:2605.20316v1 Announce Type: new Abstract: Modern text-to-image diffusion models encode rich visual priors, but expose them only through one-way text-conditioned generation. Existing unified visi

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

Gated Normalization Removal and Scale Anchoring in Pre-Norm Transformers

DGX agent

arXiv:2602.10408v2 Announce Type: replace-cross Abstract: Normalization layers are standard in transformers, but it is not clear whether their sample-dependent computations are necessary throughout bo

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

GenAI-Driven Threat Detection with Microsoft Security Copilot

DGX agent

arXiv:2605.20896v1 Announce Type: cross Abstract: Defending against today's increasingly sophisticated cyberattacks requires security analysts to continuously translate evolving attacker tradecraft in

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

Geometry-Lite: Interpretable Safety Probing via Layer-Wise Margin Geometry

DGX agent

arXiv:2605.20241v1 Announce Type: cross Abstract: Prompt-level safety probes for large language models use hidden-state representations to separate safe from unsafe prompts, but strong average detecti

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

Google just revealed Omni, personalized cross-device intelligence, and Spark agents at I/O 2025. I sat down with CEO Sundar Pichai to figure…

DGX agent

Google just revealed Omni, personalized cross-device intelligence, and Spark agents at I/O 2025. I sat down with CEO Sundar Pichai to figure out what comes next: 1:46 Omni: 'Nano Banana for video' 4:5

model-releasesrowan-cheung--x
21 May 2026
Model Releases

GradPower: Powering Gradients for Faster Language Model Pre-Training

DGX agent

arXiv:2505.24275v3 Announce Type: replace Abstract: We propose GradPower, a lightweight gradient-transformation technique for accelerating language model pre-training. Given a gradient vector g=(g_i)_

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

GraphRAG on Consumer Hardware: Benchmarking Local LLMs for Healthcare EHR Schema Retrieval

DGX agent

arXiv:2605.20815v1 Announce Type: new Abstract: Graph-based Retrieval Augmented Generation (GraphRAG) extends retrieval-augmented generation to support structured reasoning over complex corpora, but i

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

Group-Algebraic Tensors: Provably-optimal Equivariant Learning and Physical Symmetry Discovery

DGX agent

arXiv:2605.20440v1 Announce Type: new Abstract: We introduce the star_G tensor algebra, in which any finite group G defines the multiplication rule, making equivariance an intrinsic algebraic property

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

Hack-Verifiable Environments: Towards Evaluating Reward Hacking at Scale

DGX agent

arXiv:2605.20744v1 Announce Type: new Abstract: Aligning autonomous agents with human intent remains a central challenge in modern AI. A key manifestation of this challenge is reward hacking, whereby

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

HalluCXR: Benchmarking and Mitigating Hallucinations in Medical Vision-Language Models for Chest Radiograph Interpretation

DGX agent

arXiv:2605.20469v1 Announce Type: new Abstract: Vision-language models (VLMs) are increasingly used for medical image interpretation, yet they frequently hallucinate, generating clinically plausible b

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

How Much Online RL is Enough? Informative Rollouts for Offline Preference Optimization in RLVR

DGX agent

arXiv:2605.21266v1 Announce Type: new Abstract: Reinforcement Learning from Verifiable Rewards (RLVR) has emerged as a powerful paradigm for reasoning in language models, with GRPO as its primary exam

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

HRM-Text: Efficient Pretraining Beyond Scaling

DGX agent

arXiv:2605.20613v1 Announce Type: new Abstract: The current pretraining paradigm for large language models relies on massive compute and internet-scale raw text, creating a significant barrier to foun

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

Hugging Face just released LeRobot Humanoid An open-source, low-cost (~$2.5k), 3D-printed humanoid built for robot learning and not just dem…

DGX agent

Hugging Face just released LeRobot Humanoid An open-source, low-cost (~$2.5k), 3D-printed humanoid built for robot learning and not just demos. What’s cool is it’s a full stack release: • hardware + C

model-releasesclem-delangue--x
21 May 2026
Model Releases

@huggingface's @LeRobotHF team announced LeRobot Humanoid, an open-source bipedal robot platform built for roughly $2,500 using mostly 3D-pr…

DGX agent

@huggingface's @LeRobotHF team announced LeRobot Humanoid, an open-source bipedal robot platform built for roughly $2,500 using mostly 3D-printed and off-the-shelf parts: The release provides complete

model-releasesclem-delangue--x
21 May 2026
Model Releases

Humanoid Whole-Body Manipulation via Active Spatial Brain and Generalizable Action Cerebellum

DGX agent

arXiv:2605.21133v1 Announce Type: new Abstract: In this paper, we explore spatial-aware humanoid whole-body manipulation task. Compared with tabletop settings, this task poses two key challenges: 1) S

model-releasesarxiv-cs-ro
21 May 2026
Model Releases

Hyper-V2X: Hypernetworks for Estimating Epistemic and Aleatoric Uncertainty in Cooperative Bird's-Eye-View Semantic Segmentation

DGX agent

arXiv:2605.21309v1 Announce Type: new Abstract: Cooperative perception enabled by Vehicle-to-Everything (V2X) communication enhances autonomous driving safety by creating a unified environmental repre

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

I released the first alpha of Datasette Agent - a conversational AI assistant for Datasette that can answer questions about data in SQLite d…

DGX agent

I released the first alpha of Datasette Agent - a conversational AI assistant for Datasette that can answer questions about data in SQLite databases, and can be extended with plugins to add extra tool

model-releasessimon-willison--x
21 May 2026
Model Releases

If this is true, using the best public estimates we have of LLM resource use, solving this Erdos problem took 0.6–6.3 kWh of electricity and…

DGX agent

If this is true, using the best public estimates we have of LLM resource use, solving this Erdos problem took 0.6–6.3 kWh of electricity and about 3–31 liters of water. So that is less than three almo

model-releasesethan-mollick--x
21 May 2026
Model Releases

Improving Quantized Model Performance in Qualitative Analysis with Multi-Pass Prompt Verification

DGX agent

arXiv:2605.20193v1 Announce Type: new Abstract: Quantized Large Language Models (LLMs) are used more often in qualitative analysis because they run fast and need fewer computing resources. This study

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

In the next version of Claude Code: run /usage to see a breakdown of which Skills, Agents, MCPs, and Plugins are using your tokens CLI today…

DGX agent

The next version of Claude Code will introduce a `/usage` command that provides a detailed breakdown of token consumption across different components including Skills, Agents, MCPs (Model Context Prot

model-releasesboris-cherny--x
21 May 2026
Model Releases

InternBootcamp Technical Report: Boosting LLM Reasoning with Verifiable Task Scaling

DGX agent

arXiv:2508.08636v2 Announce Type: replace Abstract: Large language models (LLMs) have revolutionized artificial intelligence by enabling complex reasoning capabilities. While recent advancements in re

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

JFAA: Technical Report for the EPIC-KITCHENS-100 Action Anticipation Challenge at EgoVis 2026

DGX agent

arXiv:2605.20904v1 Announce Type: new Abstract: We propose JFAA, a JEPA-based Future Action Anticipation method for the EPIC-KITCHENS-100 (EK-100) Action Anticipation task. Inspired by the representat

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

JobArabi: An Arabic Corpus and Analysis of Job Announcements from Social Media

DGX agent

arXiv:2605.20960v1 Announce Type: new Abstract: This paper introduces JobArabi, a large-scale corpus of Arabic job announcements collected from social media between January 2024 and October 2025. The

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

JUDO: A Juxtaposed Domain-Oriented Multimodal Reasoner for Industrial Anomaly QA

DGX agent

arXiv:2605.20284v1 Announce Type: new Abstract: Industrial anomaly detection has been significantly advanced by Large Multimodal Models (LMMs), enabling diverse human instructions beyond detection, pa

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

LAION-C: An Out-of-Distribution Benchmark for Web-Scale Vision Models

DGX agent

arXiv:2506.16950v2 Announce Type: replace Abstract: Out-of-distribution (OOD) robustness is a desired property of computer vision models. Improving model robustness requires high-quality signals from

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

LamPO: A Lambda Style Policy Optimization for Reasoning Language Models

DGX agent

arXiv:2605.21235v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has become an effective paradigm for improving reasoning language models on tasks such as mathemat

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

Large-Step Training Dynamics of a Two-Factor Linear Transformer Model

DGX agent

arXiv:2605.21292v1 Announce Type: cross Abstract: Gradient-flow analyses show that simplified linear transformers can learn the in-context linear-regression algorithm, but they do not explain the fini

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

Lean Refactor: Multi-Objective Controllable Proof Optimization via Agentic Strategy Search

DGX agent

arXiv:2605.20244v1 Announce Type: cross Abstract: We present Lean Refactor, a plug-and-play retrieval-augmented agentic framework for multi-objective, controllable, and version-robust refactoring of L

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

LEAP: A closed-loop framework for perovskite precursor additive discovery

DGX agent

arXiv:2605.20242v1 Announce Type: new Abstract: Efficient discovery of precursor additives is essential for improving the performance of perovskite solar cells, yet the large chemical space makes conv

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

Learning fMRI activations dictionaries across individual geometries via optimal transport

DGX agent

arXiv:2605.20883v1 Announce Type: new Abstract: Dictionary learning is a powerful tool for creating interpretable representations. When applied to functional magnetic resonance imaging (fMRI) data, th

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

Learning Query-Aware Budget-Tier Routing for Runtime Agent Memory

DGX agent

arXiv:2602.06025v2 Announce Type: replace Abstract: Memory is increasingly central to Large Language Model (LLM) agents operating beyond a single context window, yet most existing systems rely on offl

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

Learning-to-Defer with Expert-Conditional Advice

DGX agent

arXiv:2603.14324v3 Announce Type: replace-cross Abstract: Learning-to-Defer routes each input to the expert that minimizes expected cost, but it assumes that the information available to every expert

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

LER-YOLO: Reliability-Aware Expert Routing for Misaligned RGB-Infrared UAV Detection

DGX agent

arXiv:2605.20667v1 Announce Type: new Abstract: Detecting small unmanned aerial vehicles from RGB-infrared remote-sensing pairs remains challenging due to tiny target scale, cluttered backgrounds, and

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

Leveraging LLMs for Grammar Adaptation: A Study on Metamodel-Grammar Co-Evolution

DGX agent

arXiv:2605.21465v1 Announce Type: new Abstract: In model-driven engineering, metamodel evolution leads to the need to adapt corresponding grammars to maintain consistency, which typically requires ted

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

Leveraging Vision-Language Models to Detect Attention in Educational Videos

DGX agent

arXiv:2605.20211v1 Announce Type: new Abstract: Educational videos are a cornerstone of remote and blended learning. However, learners' fluctuating attention remains a significant barrier to effective

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

Llamas on the Web: Memory-Efficient, Performance-Portable, and Multi-Precision LLM Inference with WebGPU

DGX agent

arXiv:2605.20706v1 Announce Type: cross Abstract: Running language models in the browser presents a unique opportunity to build efficient, private, and portable AI applications, but requires contendin

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

LLMs on the Line: Data Determines Loss-to-Loss Scaling Laws

DGX agent

arXiv:2502.12120v3 Announce Type: replace-cross Abstract: Scaling laws guide the development of large language models (LLMs) by offering estimates for the optimal balance of model size, tokens, and co

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

MagBridge-Battery: A Synthetic Bridge Dataset for Li-ion Magnetometry and State-of-Health Diagnostics

DGX agent

arXiv:2605.20240v1 Announce Type: new Abstract: Battery health diagnostics today rely overwhelmingly on electrochemical signals measured at the cell terminals. A parallel literature has shown that mag

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

Markovian Circuit Tracing for Transformer State Dynamic

DGX agent

arXiv:2605.20824v1 Announce Type: new Abstract: Many sequence computations are easier to study as movement through internal states than as isolated local circuits. We introduce Markovian Circuit Traci

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

Mechanics of Bias and Reasoning: Interpreting the Impact of Chain-of-Thought Prompting on Gender Bias in LLMs

DGX agent

arXiv:2605.20410v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed in socially sensitive settings despite substantial documentation that they encode gender biases.

model-releasesarxiv-cs-cl
21 May 2026
← Previous
1…281282283284285…472
Next →