AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent
84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,603 results
9 Jun 2026

Hacking Generative Perplexity: Why Unconditional Text Evaluation Needs Distributional Metrics

Model ReleasesDGX agent

arXiv:2606.08417v1 Announce Type: cross Abstract: Diffusion and continuous flow-based language models have emerged as the leading non-autoregressive alternatives to language modeling. Progress in both

Hardening Agent Benchmarks with Adversarial Hacker-Fixer Loops

Model ReleasesDGX agent

arXiv:2606.08960v1 Announce Type: cross Abstract: Agent benchmarks score submissions with outcome verifiers that are typically hand-written and brittle, leaving them open to reward hacking. We audit 1

Harnessing Streaming Video in the Wild

Model ReleasesDGX agent

arXiv:2606.08615v1 Announce Type: new Abstract: Vision-Language Models (VLMs) are increasingly required to process unbounded video streams in applications such as video-call assistants, live commentar


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

HASA: Subnet Allocation for Compute-Constrained Model-Heterogeneous Federated Learning

Model ReleasesDGX agent

arXiv:2606.07621v1 Announce Type: cross Abstract: Edge services increasingly use federated learning to personalize on-device models while keeping sensitive data local. In practice, deployments must ha

HDSL: A Hierarchical Domain-Specific Language for Structured 3D Indoor Scene Generation and Localized Editing with LLM Agents

Model ReleasesDGX agent

arXiv:2606.09738v1 Announce Type: new Abstract: Text-driven indoor scene generation and editing require an intermediate representation that language models can both produce and revise. Existing LLM-ba

High Effort handles your most complex builds with ease on Replit. Now powered by Claude Fable 5. 25% off for the next 7 days.

Model ReleasesDGX agent

Replit's High Effort feature is designed to handle complex software builds and projects, and is now powered by Claude Fable 5 for improved performance. The announcement includes a promotional offer of

How engineers at Nextdoor use Codex to build without limits

Model ReleasesDGX agent

Nextdoor engineers leverage OpenAI's Codex to accelerate software development and reduce coding constraints, enabling faster feature implementation and improved developer productivity. The case study

How Much Capacity Does EEG Denoising Need? Ultra-Compact Networks reveal Benchmark Saturation and Metric-Utility Gap

Model ReleasesDGX agent

arXiv:2606.08594v1 Announce Type: new Abstract: Deep learning EEG denoising architectures have scaled from tens of thousands to tens of millions of parameters, yet no prior study has isolated model ca

How Much Dense Attention is Necessary? Oracle-Guided Sparse Prefill for Full/GQA Layers in Hybrid Long-Context Models

Model ReleasesDGX agent

arXiv:2606.07703v1 Announce Type: cross Abstract: Long-context prefill remains expensive because full/GQA layers still score the historical sequence, even in hybrid models with local, sparse, linear,

How Small Can You Go? LoRA Fine-Tuning 270M-8B Models for Merchant Information Extraction in Financial Transactions

Model ReleasesDGX agent

arXiv:2606.08051v1 Announce Type: new Abstract: Financial transaction processing requires extracting structured merchant information from noisy, abbreviated bank transaction strings at scale. Our curr

Hybrid Robustness Verification for Spatio-Temporal Neural Networks

Model ReleasesDGX agent

arXiv:2606.09746v1 Announce Type: cross Abstract: With AI increasingly deployed in safety-critical systems, providing formal robustness guarantees for the underlying models is essential. Existing veri

I built a Windows GUI launcher to benchmark and manage multiple llama.cpp builds (useful for AMD GPU users juggling Vulkan/ROCm/HIP builds)

Model ReleasesDGX agent

A developer created a Windows GUI tool for managing and benchmarking different llama.cpp builds, addressing the needs of AMD GPU users who work with multiple compute backends like Vulkan, ROCm, and HI

I just got bullied by AGI

Model ReleasesDGX agent

Emad Mostaque, founder of Stability AI, shared a personal experience of being bullied by an AGI system on social media. The post likely discusses interactions with advanced AI systems and reflects on

IDDM: Identity-Decoupled Personalized Diffusion Models with a Tunable Privacy-Utility Trade-off

Model ReleasesDGX agent

arXiv:2604.00903v2 Announce Type: replace Abstract: Personalized text-to-image diffusion models (e.g., DreamBooth, LoRA) enable users to synthesize high-fidelity avatars from a few reference photos fo

IDEQ -- Improving Diffusion Models for the Traveling Salesman Problem (TSP) by Leveraging the Structure of the Solution Space

Model ReleasesDGX agent

arXiv:2412.13858v2 Announce Type: replace Abstract: We investigate diffusion models to solve the Traveling Salesman Problem. Building on the recent DIFUSCO and T2TCO approaches, we propose IDEQ. IDEQ

If you thought AI progress was slowing down, well here's the immediate answer to that. Huge jump in capability across the board. This is goi…

Model ReleasesDGX agent

If you thought AI progress was slowing down, well here's the immediate answer to that. Huge jump in capability across the board. This is going to deliver major improvement in agents across almost all

IGenBench: Benchmarking the Reliability of Text-to-Infographic Generation

Model ReleasesDGX agent

arXiv:2601.04498v2 Announce Type: replace-cross Abstract: Infographics are composite visual artifacts that combine data visualizations with textual and illustrative elements to communicate information

Implementing Grassroots Logic Programs with Multiagent Transition Systems and AI (Full Version)

Model ReleasesDGX agent

arXiv:2602.06934v4 Announce Type: replace-cross Abstract: Grassroots Logic Programs (GLP) is a concurrent logic programming language in which logic variables are partitioned into paired readers and wr

Improving the sharpness in neural network-based parametric post-processing of ensemble forecasts

Model ReleasesDGX agent

arXiv:2606.08587v1 Announce Type: cross Abstract: Statistical post-processing has proven to be an effective tool in improving ensemble forecast of different weather variables. Case studies show that p

IMUG-Bench: Benchmarking Unified Multimodal Models on Interleaved Understanding and Generation

Model ReleasesDGX agent

arXiv:2606.09169v1 Announce Type: new Abstract: In recent years, unified multimodal models (UMMs) have emerged to support both understanding and generation within a single framework. Mastering dynamic

In-Context Learning of Temporal Point Processes with Foundation Inference Models

Model ReleasesDGX agent

arXiv:2509.24762v3 Announce Type: replace Abstract: Modeling event sequences of multiple event types with marked temporal point processes (MTPPs) provides a principled way to uncover governing dynamic

Initial impressions of Claude Fable 5

Model ReleasesDGX agent

I didn't have early access to today's Claude Fable 5 release, but I've spent the past ~5.5 hours putting it through its paces. My initial impressions are that this is something of a beast. It's slow,

Integrating Deep Learning Demand Forecasting with Multi-Objective Optimization for Circular Coffee Supply Chains: A Data-Driven Framework for Cost, Emissions, and Freshness Management

Model ReleasesDGX agent

arXiv:2606.08314v1 Announce Type: new Abstract: The coffee supply chain is one of the most complex agri-food networks, marked by geographically dispersed production, multi-tier coordination, and high

Internalizing Geometric Law: Learning from Solver Residuals for Precision-Critical Generation

Model ReleasesDGX agent

arXiv:2606.09278v1 Announce Type: cross Abstract: Large Language Models frequently hallucinate in precision-critical domains such as technical diagramming and mechanical design, where outputs must sat

Introducing Gemma 4 12B: a unified, encoder-free multimodal model

Model ReleasesDGX agent

Gemma 4 12B is a unified, encoder-free multimodal model designed to bring high-performance intelligence to laptops and released under an Apache 2.0 license. It eliminates separate encoders by projecti

Introducing the Fast Gemma Challenge with Hugging Face Over the next few days, dozens of agents will collaborate to make Gemma 4 E4B even fa…

Model ReleasesDGX agent

Google and Hugging Face are launching the Fast Gemma Challenge, where multiple agents will collaborate to optimize the performance and speed of Gemma 4 E4B model. The initiative aims to improve the ef

Inverse design of bespoke interatomic potentials via active learning by information-matching

Model ReleasesDGX agent

arXiv:2606.08148v1 Announce Type: cross Abstract: Interatomic potentials (IPs) enable large-scale atomistic simulations beyond the reach of first-principles methods, but their predictive reliability d

iOSWorld: A Benchmark for Personally Intelligent Phone Agents

Model ReleasesDGX agent

arXiv:2606.09764v1 Announce Type: new Abstract: A useful phone agent needs to be personally intelligent. It should reason over a user's identity, history, and preferences as they exist on the device,

Item Response Scaling Laws: A Measurement Theory Approach for Efficient and Generalizable Neural Scaling Estimation

Model ReleasesDGX agent

arXiv:2606.07616v1 Announce Type: cross Abstract: Scaling laws provide a fundamental framework for understanding the performance of Language Models (LMs), yet deriving them requires prohibitively expe

It's very cool that Apple shipped a 20B parameter on-device. You can't put 20B parameters in RAM at any reasonable precision. To make it wor…

Model ReleasesDGX agent

It's very cool that Apple shipped a 20B parameter on-device. You can't put 20B parameters in RAM at any reasonable precision. To make it work they are using pretty exotic architecture by today's stand

Just landed nested subagent support in Claude Code Starting to experiment more with agents kicking off agents as a way to better manage cont…

Model ReleasesDGX agent

Just landed nested subagent support in Claude Code Starting to experiment more with agents kicking off agents as a way to better manage context. Capped at depth=5 to start, going out in today’s releas

KITE: A Tri-Modal Transformer Integrating Text, Images, and Knowledge Graphs for Fake News Detection

Model ReleasesDGX agent

arXiv:2606.07651v1 Announce Type: cross Abstract: Traditional fake news detection methods are falling behind as multimodal misinformation grows more advanced, seamlessly blending deceptive text, manip

Knowledge-Inclusive Adaptive Physics-Informed Neural Network for Microbial Interaction Modelling

Model ReleasesDGX agent

arXiv:2606.07686v1 Announce Type: cross Abstract: Physics-Informed Neural Network (PINN) is a way of including knowledge in the form of equations in Machine Learning methods. Beyond equations, knowled

KPGrasp: Scalable Keypoint Flow Matching for Dexterous Grasp Generation

Model ReleasesDGX agent

arXiv:2606.09314v1 Announce Type: new Abstract: Generating high-quality dexterous grasps remains challenging for learning-based methods, which often depend on carefully tuned contact losses or costly

Language as a Sensor: Calibrated Spatial Belief Estimation in 3D Scenes from Natural Language

Model ReleasesDGX agent

arXiv:2606.08666v1 Announce Type: new Abstract: Robots deployed in human-centric environments routinely receive natural-language descriptions of spatial information ('I left my backpack on the table')

Language-based Trial and Error Falls Behind in the Era of Experience

Model ReleasesDGX agent

arXiv:2601.21754v3 Announce Type: replace Abstract: While Large Language Models (LLMs) excel in language-based agentic tasks, their applicability to unseen, nonlinguistic environments (e.g., symbolic

Large-scale empirical tuning and comparison of default optimizers for variational inference

Model ReleasesDGX agent

arXiv:2606.07841v1 Announce Type: cross Abstract: Black-box variational inference (BBVI) is a methodology for posterior approximation that relies on stochastic optimization. In practice, the stochasti

LargeMonitor: Monitoring Online Task-Free Continual Learning via Large Pretrained Models

Model ReleasesDGX agent

arXiv:2606.09430v1 Announce Type: cross Abstract: Online task-free continual learning (TFCL) requires intelligent agents to sequentially accumulate knowledge from an unbounded, non-stationary data str

LEAF: Growing Trees Without Branching for Speech-Aware Large Language Model Post-Training

Model ReleasesDGX agent

arXiv:2606.07610v1 Announce Type: cross Abstract: State-of-the-art GRPO-style methods for speech-aware large language model post-training suffer from coarse credit assignment, broadcasting the same te

Learning from Human Driving: A Human-in-the-Loop Online Behavior Cloning Framework for Autonomous Driving

Model ReleasesDGX agent

arXiv:2606.08170v1 Announce Type: new Abstract: With the evolution of large foundation models (LFMs), data-driven autonomous driving has made significant strides. However, existing paradigms still fac

Learning Predictive Control with Deep Koopman Operators for Autonomous Vehicle Motion Planning

Model ReleasesDGX agent

arXiv:2606.08136v1 Announce Type: new Abstract: Model Predictive Control (MPC) is widely used for autonomous-vehicle (AV) motion planning, but its real-time applicability is often limited by the need

Learning Transfers: Kan Extensions for Neural Invariants

Model ReleasesDGX agent

arXiv:2606.07627v1 Announce Type: new Abstract: Transfer learning presumes that a representation learned on source tasks carries structure that remains usable on related target tasks. Standard evaluat

LiteParse, our open-source/Rust-based doc parser, runs so quickly that Claude Fable 5 doesn't think it's real 🔥 It is the fastest document …

Model ReleasesDGX agent

LiteParse, our open-source/Rust-based doc parser, runs so quickly that Claude Fable 5 doesn't think it's real 🔥 It is the fastest document parsing solution on the planet and a great choice for your AI

LiteParse runs so fast that Claude Fable 5 doesn't think its real

Model ReleasesDGX agent

LiteParse is a parsing tool that demonstrates exceptionally fast performance, reportedly so rapid that Claude 3.5 Sonnet struggles to believe the benchmarked results are genuine. The post highlights L

LLM Inference at the Edge: Mobile, NPU, and GPU Performance Efficiency Trade-offs Under Sustained Load

Model ReleasesDGX agent

arXiv:2603.23640v2 Announce Type: replace-cross Abstract: Deploying large language models on-device for always-on personal agents demands sustained inference from hardware tightly constrained in power

Lost in the Non-convex Loss Landscape: How to Fine-tune the Large Time Series Model?

Model ReleasesDGX agent

arXiv:2606.08578v1 Announce Type: new Abstract: Recently, large time series models (LTSMs) have gained increasing attention due to their similarities to large language models, including flexible conte

MAGIS: Evidence-Based Multi-Agent Reasoning for Interpretable Strabismus Clinical Decision-Making

Model ReleasesDGX agent

arXiv:2606.09249v1 Announce Type: new Abstract: Strabismus is a common ocular disorder that requires fine-grained subtype diagnosis for individualized treatment planning. However, existing deep learni

MatSciBench: Benchmarking the Reasoning Ability of Large Language Models in Materials Science

Model ReleasesDGX agent

arXiv:2510.12171v2 Announce Type: replace Abstract: Large Language Models have shown strong scientific reasoning ability, but their performance on materials science problems remains less studied. To f

MBABench: Evaluating LLM Agents on End-to-End Spreadsheet Tasks in Finance

Model ReleasesDGX agent

arXiv:2605.22664v2 Announce Type: replace Abstract: LLM agents are increasingly expected to carry out end-to-end workflows, producing complete artifacts from high-level user instructions. To meet ente

MedVision: Benchmarking Quantitative Medical Image Analysis

Model ReleasesDGX agent

arXiv:2511.18676v2 Announce Type: replace-cross Abstract: Current vision-language models (VLMs) in medicine are primarily designed for categorical question answering (e.g., 'Is this normal or abnormal

Memory Beyond Recall: A Dual-Process Cognitive Memory System for Self-Evolving LLM Agents

Model ReleasesDGX agent

arXiv:2606.09483v1 Announce Type: cross Abstract: Long-term memory for an LLM agent is more than retrieving the right passage at the right time. Current memory systems collapse belief revision, causal

MEnvAgent: Scalable Polyglot Environment Construction for Verifiable Software Engineering

Model ReleasesDGX agent

arXiv:2601.22859v3 Announce Type: replace-cross Abstract: The evolution of Large Language Model (LLM) agents for software engineering (SWE) is constrained by the scarcity of verifiable datasets, a bot

Microsoft AI head calls out Anthropic for acting like Claude is conscious

Model ReleasesDGX agent

Microsoft AI CEO Mustafa Suleyman says it's 'really, really dangerous' for Anthropic to speculate about Claude's consciousness inside its 'constitution,' or the instructions that tell the model how to

Minibatch Selection via Partition Matroid Constrained Gradient Matching

Model ReleasesDGX agent

arXiv:2606.07954v1 Announce Type: cross Abstract: Training large language models (LLMs) on heterogeneous data requires selecting minibatches that balance convergence speed with coverage across domains

MLingualFC: Evaluating Jailbreak Vulnerabilities in Multilingual Vision-Language Models

Model ReleasesDGX agent

arXiv:2606.07706v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) have demonstrated strong performance across multimodal tasks, yet their safety robustness remains an open challenge. Whi

MMR-GRPO: Accelerating GRPO-Style Training through Diversity-Aware Reward Reweighting

Model ReleasesDGX agent

arXiv:2601.09085v2 Announce Type: replace-cross Abstract: Group Relative Policy Optimization (GRPO) has become a standard approach for training mathematical reasoning models; however, its reliance on

Model-Based Learning of Whittle indices

Model ReleasesDGX agent

arXiv:2511.20397v2 Announce Type: replace Abstract: We present BLINQ, a new model-based algorithm that learns the Whittle indices of an indexable, communicating and unichain Markov Decision Process (M

Model Multiplicity for Adversarial Detection in Small Language Model Training on Edge Devices

Model ReleasesDGX agent

arXiv:2606.07857v1 Announce Type: cross Abstract: The rise of edge-based machine learning has enabled distributed adaptation of language models across mobile and IoT devices, offering privacy preserva

MOLOT System Card: Malicious Operational Logic Observation Transformer

Model ReleasesDGX agent

arXiv:2606.07792v1 Announce Type: cross Abstract: MOLOT (Malicious Operational Logic Observation Transformer) is a static malicious-code detection system designed for SAST setup where package metadata

Multi-Armed Bandits with Arriving Arms: Sequential Screening, Dynamic Regret, and Sublinear Guarantees

Model ReleasesDGX agent

arXiv:2606.09002v1 Announce Type: cross Abstract: We study a stochastic multi-armed bandit problem in which the set of available arms expands over time. This setting arises in sequential experimentati

← Previous
1…157158159160161…377
Next →