AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
All
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,569 results
Model Releases

PERSA: Reinforcement Learning for Professor-Style Personalized Feedback with LLMs

DGX agent

arXiv:2605.01123v1 Announce Type: new Abstract: Large language models (LLMs) can provide automated feedback in educational settings, but aligning an LLMs style with a specific instructors tone while m

model-releasesarxiv-cs-ai
6 May 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

PHBench: A Benchmark for Predicting Startup Series A Funding from Product Hunt Launch Signals

DGX agent

arXiv:2605.02974v1 Announce Type: cross Abstract: Structured launch signals on Product Hunt contain statistically significant predictive information for Series A funding outcomes. We construct PHBench

model-releasesarxiv-cs-lg
6 May 2026
Model Releases

PhysicianBench: Evaluating LLM Agents in Real-World EHR Environments

DGX agent

arXiv:2605.02240v1 Announce Type: new Abstract: We introduce PhysicianBench, a benchmark for evaluating LLM agents on physician tasks grounded in real clinical setting within electronic health record

model-releasesarxiv-cs-ai
6 May 2026
Model Releases

PIIGuard: Mitigating PII Harvesting under Adversarial Sanitization

DGX agent

arXiv:2605.03129v1 Announce Type: cross Abstract: Browsing-enabled LLM assistants can fetch webpages and answer contact-seeking queries, creating a practical channel for scraping contact-style persona

model-releasesarxiv-cs-cl
6 May 2026
Model Releases

Pioneering AI-assisted code migration: How Google achieved 6x faster migration from TensorFlow to JAX

DGX agent

AI coding agents are rapidly becoming ubiquitous across the software industry, fundamentally changing how developers write, test, and debug daily code. While these tools excel at localized, self-conta

model-releasesgoogle-cloud-ai
6 May 2026
Model Releases

PODiff: Latent Diffusion in Proper Orthogonal Decomposition Space for Scientific Super-Resolution

DGX agent

arXiv:2605.03399v1 Announce Type: new Abstract: Probabilistic super-resolution of high-dimensional spatial fields using diffusion models is often computationally prohibitive due to the cost of operati

model-releasesarxiv-cs-lg
6 May 2026
Model Releases

Preview in Claude Code on desktop makes it easy to give Claude context. You can attach DOM elements or use the pencil tool to draw directly …

DGX agent

Claude Code on desktop includes a preview feature that allows users to provide context to Claude by attaching DOM elements or drawing directly with a pencil tool. This functionality streamlines the pr

model-releasesthariq--x
6 May 2026
Model Releases

PriorNet: Prior-Guided Engagement Estimation from Face Video

DGX agent

arXiv:2605.03615v1 Announce Type: new Abstract: Engagement estimation from face video remains challenging because facial evidence is often incomplete, labeled data are limited, and engagement annotati

model-releasesarxiv-cs-cv
6 May 2026
Model Releases

Prism: Efficient Test-Time Scaling via Hierarchical Search and Self-Verification for Discrete Diffusion Language Models

DGX agent

arXiv:2602.01842v3 Announce Type: replace Abstract: Inference-time compute has re-emerged as a practical way to improve LLM reasoning. Most test-time scaling (TTS) algorithms rely on autoregressive de

model-releasesarxiv-cs-lg
6 May 2026
Model Releases

PROBE: Probabilistic Occupancy BEV Encoding with Analytical Translation Robustness for 3D Place Recognition

DGX agent

arXiv:2603.05965v2 Announce Type: replace-cross Abstract: We present PROBE (PRobabilistic Occupancy BEV Encoding), a learning-free LiDAR place recognition descriptor that models each BEV cell's occupa

model-releasesarxiv-cs-cv
6 May 2026
Model Releases

Quantum Hierarchical Reinforcement Learning via Variational Quantum Circuits

DGX agent

arXiv:2605.03434v1 Announce Type: new Abstract: Reinforcement learning is one of the most challenging learning paradigms where efficacy and efficiency gains are extremely valuable. Hierarchical reinfo

model-releasesarxiv-cs-lg
6 May 2026
Model Releases

RAG over Thinking Traces Can Improve Reasoning Tasks

DGX agent

arXiv:2605.03344v1 Announce Type: cross Abstract: Retrieval-augmented generation (RAG) has proven effective for knowledge-intensive tasks, but is widely believed to offer limited benefit for reasoning

model-releasesarxiv-cs-cl
6 May 2026
Model Releases

Raising the Ceiling: Better Empirical Fixation Densities for Saliency Benchmarking

DGX agent

arXiv:2605.03885v1 Announce Type: new Abstract: Empirical fixation densities, spatial distributions estimated from human eye-tracking data, are foundational to saliency benchmarking. They directly sha

model-releasesarxiv-cs-cv
6 May 2026
Model Releases

RD-ViT: Recurrent-Depth Vision Transformer for Semantic Segmentation with Reduced Data Dependence Extending the Recurrent-Depth Transformer Architecture to Dense Prediction

DGX agent

arXiv:2605.03999v1 Announce Type: new Abstract: Vision Transformers (ViTs) achieve state-of-the-art segmentation accuracy but require large training datasets because each layer has unique parameters t

model-releasesarxiv-cs-cv
6 May 2026
Model Releases

Real Image Denoising with Knowledge Distillation for High-Performance Mobile NPUs

DGX agent

arXiv:2605.03680v1 Announce Type: new Abstract: While deep-learning-based image restoration has achieved unprecedented fidelity, deployment on mobile Neural Processing Units (NPUs) remains bottlenecke

model-releasesarxiv-cs-cv
6 May 2026
Model Releases

Reasoning Models Can be Accurately Pruned Via Chain-of-Thought Reconstruction

DGX agent

arXiv:2509.12464v2 Announce Type: replace Abstract: Reasoning language models such as DeepSeek-R1 produce long chain-of-thought traces during inference time which make them costly to deploy at scale.

model-releasesarxiv-cs-ai
6 May 2026
Model Releases

ReCode: Reinforcing Code Generation with Reasoning-Process Rewards

DGX agent

arXiv:2508.05170v3 Announce Type: replace-cross Abstract: In practice, rigorous reasoning is often a key driver of correct code, while Reinforcement Learning (RL) for code generation often neglects op

model-releasesarxiv-cs-cl
6 May 2026
Model Releases

ReLeaf: Benchmarking Leaf Segmentation across Domains and Species

DGX agent

arXiv:2605.03784v1 Announce Type: new Abstract: Rising global food demand and growing climate pressure increase the need for sustainable, precise agricultural practices. Automated, individualized plan

model-releasesarxiv-cs-cv
6 May 2026
Model Releases

Reproducing Complex Set-Compositional Information Retrieval

DGX agent

arXiv:2605.03824v1 Announce Type: new Abstract: Complex information needs may involve set-compositional queries using conjunction, disjunction, and exclusion, yet it remains unclear whether current re

model-releasesarxiv-cs-cl
6 May 2026
Model Releases

Rethinking Reasoning-Intensive Retrieval: Evaluating and Advancing Retrievers in Agentic Search Systems

DGX agent

arXiv:2605.04018v1 Announce Type: new Abstract: Reasoning-intensive retrieval aims to surface evidence that supports downstream reasoning rather than merely matching topical similarity. This capabilit

model-releasesarxiv-cs-cl
6 May 2026
Model Releases

Retrieval and Multi-Hop Reasoning in 1M-Token Context Windows: Evaluating LLMs on Classical Chinese Text

DGX agent

arXiv:2605.02173v1 Announce Type: new Abstract: We evaluate the long-context retrieval and reasoning capabilities of five frontier large language models with advertised 1M-token context windows on a c

model-releasesarxiv-cs-ai
6 May 2026
Model Releases

Retrieval-Augmented LLMs for Security Incident Analysis

DGX agent

arXiv:2603.18196v3 Announce Type: replace-cross Abstract: Investigating cybersecurity incidents requires collecting and analyzing evidence from multiple log sources, including intrusion detection aler

model-releasesarxiv-cs-ai
6 May 2026
Model Releases

Reward Hacking Benchmark: Measuring Exploits in LLM Agents with Tool Use

DGX agent

arXiv:2605.02964v1 Announce Type: new Abstract: Reinforcement learning (RL) trained language model agents with tool access are increasingly deployed in coding assistants, research tools, and autonomou

model-releasesarxiv-cs-lg
6 May 2026
Model Releases

RFPrompt: Prompt-Based Expert Adaptation of the Large Wireless Model for Modulation Classification

DGX agent

arXiv:2605.03279v1 Announce Type: new Abstract: Automatic modulation classification (AMC) in real-world deployments demands robustness to distribution shifts arising from hardware impairments, unseen

model-releasesarxiv-cs-lg
6 May 2026
Model Releases

RoboAlign-R1: Distilled Multimodal Reward Alignment for Robot Video World Models

DGX agent

arXiv:2605.03821v1 Announce Type: new Abstract: Existing robot video world models are typically trained with low-level objectives such as reconstruction and perceptual similarity, which are poorly ali

model-releasesarxiv-cs-ro
6 May 2026
Model Releases

RoboEval: Where Robotic Manipulation Meets Structured and Scalable Evaluation

DGX agent

arXiv:2507.00435v2 Announce Type: replace-cross Abstract: We introduce RoboEval, a structured evaluation framework and benchmark for robotic manipulation that augments binary success with principled b

model-releasesarxiv-cs-cv
6 May 2026
Model Releases

Robust Language Identification for Romansh Varieties

DGX agent

arXiv:2603.15969v2 Announce Type: replace Abstract: The Romansh language has several regional varieties, called idioms, which sometimes have limited mutual intelligibility. Despite this linguistic div

model-releasesarxiv-cs-cl
6 May 2026
Model Releases

S2O: Early Stopping for Sparse Attention via Online Permutation

DGX agent

arXiv:2602.22575v2 Announce Type: replace Abstract: Attention scales quadratically with sequence length, fundamentally limiting long-context inference. Existing block-granularity sparsification can re

model-releasesarxiv-cs-lg
6 May 2026
Model Releases

Safety and accuracy follow different scaling laws in clinical large language models

DGX agent

arXiv:2605.04039v1 Announce Type: new Abstract: Clinical LLMs are often scaled by increasing model size, context length, retrieval complexity, or inference-time compute, with the implicit expectation

model-releasesarxiv-cs-cl
6 May 2026
Model Releases

SAM-NER: Semantic Archetype Mediation for Zero-Shot Named Entity Recognition

DGX agent

arXiv:2605.03706v1 Announce Type: new Abstract: Zero-shot Named Entity Recognition (ZS-NER) remains brittle under domain and schema shifts, where unseen label definitions often misalign with a large l

model-releasesarxiv-cs-cl
6 May 2026
Model Releases

Scaling Unsupervised Multi-Source Federated Domain Adaptation through Group-Wise Discrepancy Minimization

DGX agent

arXiv:2510.08150v3 Announce Type: replace Abstract: Unsupervised multi-source domain adaptation (UMDA) leverages labeled data from multiple source domains to generalize to an unlabeled target. While f

model-releasesarxiv-cs-lg
6 May 2026
Model Releases

SCGNN: Semantic Consistency enhanced Graph Neural Network Guided by Granular-ball Computing

DGX agent

arXiv:2605.02617v2 Announce Type: new Abstract: Capturing semantic consistency among nodes is crucial for effective graph representation learning. Existing approaches typically rely on k-nearest neigh

model-releasesarxiv-cs-ai
6 May 2026
Model Releases

See what these young builders are doing with AI now that everyone can just build things. https://chatgpt.com/futures/

DGX agent

This OpenAI social media post highlights how young builders and developers are leveraging AI tools to create projects and applications now that AI development has become more accessible to non-experts

model-releasesopenai--x
6 May 2026
Model Releases

Seeking Information with RAG-Assistants: Does Model Size Matter in Human-AI Collaborations?

DGX agent

arXiv:2605.00964v1 Announce Type: cross Abstract: Much research on LLMs has focused on increasing benchmark performance. However, the evaluation of such models in real-world collaborative human-AI wor

model-releasesarxiv-cs-ai
6 May 2026
Model Releases

Self-Mined Hardness for Safety Fine-Tuning

DGX agent

arXiv:2605.03226v1 Announce Type: new Abstract: Safety fine-tuning of language models typically requires a curated adversarial dataset. We take a different approach: score each candidate prompt's diff

model-releasesarxiv-cs-lg
6 May 2026
Model Releases

Sentinel2Cap: A Human-Annotated Benchmark Dataset for Multimodal Remote Sensing Image Captioning

DGX agent

arXiv:2605.03189v1 Announce Type: new Abstract: Image captioning has become an important task in computer vision, enabling models to generate natural language descriptions of visual content. While sev

model-releasesarxiv-cs-cv
6 May 2026
Model Releases

Simulated Students in Tutoring Dialogues: Substance or Illusion?

DGX agent

arXiv:2601.04025v2 Announce Type: replace Abstract: Advances in large language models (LLMs) enable many new innovations in education. However, evaluating the effectiveness of new technology requires

model-releasesarxiv-cs-cl
6 May 2026
Model Releases

Singular Bank helps bankers move fast with ChatGPT and Codex

DGX agent

Singular Bank leverages OpenAI's ChatGPT and Codex to accelerate banking operations and developer productivity. The integration enables bankers and software engineers to streamline workflows, automate

model-releasesopenai
6 May 2026
Model Releases

so cute!

DGX agent

so cute! It's pretty magical how kids interact with tech. They just expect it to work for them. Here's mine talking to Reachy Mini via Gemini Live API. I didn't tell them how, they just did it. That's

model-releasesclem-delangue--x
6 May 2026
Model Releases

Soft Tournament Equilibrium

DGX agent

arXiv:2604.04328v3 Announce Type: replace-cross Abstract: The evaluation of general-purpose artificial agents, particularly those based on LLMs, presents a significant challenge due to the non-transit

model-releasesarxiv-cs-lg
6 May 2026
Model Releases

SpaceXAI will provide @AnthropicAI with access to Colossus 1, one of the world’s largest and fastest-deployed AI supercomputers, to provide …

DGX agent

SpaceXAI will provide @AnthropicAI with access to Colossus 1, one of the world’s largest and fastest-deployed AI supercomputers, to provide additional capacity for Claude → http://x.ai/news/anthropic-

model-releasesthariq--x
6 May 2026
Model Releases

Sparse Memory Finetuning as a Low-Forgetting Alternative to LoRA and Full Finetuning

DGX agent

arXiv:2605.03229v1 Announce Type: new Abstract: Adapting a pretrained language model to a new task often hurts the general capabilities it already had, a problem known as catastrophic forgetting. Spar

model-releasesarxiv-cs-cl
6 May 2026
Model Releases

Stable Multimodal Graph Unlearning via Feature-Dimension Aware Quantile Selection

DGX agent

arXiv:2605.03303v1 Announce Type: new Abstract: Graph unlearning remains a critical technique for supporting privacy-preserving and sustainable multimodal graph learning. However, we observe that exis

model-releasesarxiv-cs-lg
6 May 2026
Model Releases

StateSMix: Online Lossless Compression via Mamba State Space Models and Sparse N-gram Context Mixing

DGX agent

arXiv:2605.02904v1 Announce Type: new Abstract: We present StateSMix, a fully self-contained lossless compressor that couples an online-trained Mamba-style State Space Model (SSM) with sparse n-gram c

model-releasesarxiv-cs-lg
6 May 2026
Model Releases

StateVLM: A State-Aware Vision-Language Model for Robotic Affordance Reasoning

DGX agent

arXiv:2605.03927v1 Announce Type: new Abstract: Vision-language models (VLMs) have shown remarkable performance in various robotic tasks, as they can perceive visual information and understand natural

model-releasesarxiv-cs-cv
6 May 2026
Model Releases

Strong Opinions, Loosely Held on Agent + Harness Engineering: 1. You can outperform any default harness+model (including codex & claude code…

DGX agent

Strong Opinions, Loosely Held on Agent + Harness Engineering: 1. You can outperform any default harness+model (including codex & claude code) on pretty much any Task by engineering the harness around

model-releasesharrison-chase--x
6 May 2026
Model Releases

SURE-RAG: Sufficiency and Uncertainty-Aware Evidence Verification for Selective Retrieval-Augmented Generation

DGX agent

arXiv:2605.03534v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) grounds answers in retrieved passages, but retrieval is not verification: a passage can be topical and still fail t

model-releasesarxiv-cs-cl
6 May 2026
Model Releases

@swyx @maheshmurag deep dive on the new dreaming feature in beta today https://x.com/claudeai/status/2052067400690851842?s=46

DGX agent

@swyx @maheshmurag deep dive on the new dreaming feature in beta today https://x.com/claudeai/status/2052067400690851842?s=46 Dreaming reviews your agent's past sessions, extracts patterns, and curate

model-releasesswyx--x
6 May 2026
← Previous
1…354355356357358…471
Next →