AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent
84,630Total entries
1Added by human
84,629Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,603 results
9 Jun 2026

GEAR-VLA: Learning Geometry-Aware Action Representations for Generalizable Robotic Manipulation

Model ReleasesDGX agent

arXiv:2606.08530v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models achieve strong benchmark performance but still struggle in real-world deployment with unseen objects, background s

Gemini for Government: Your blueprint for mission impact

Model ReleasesDGX agent

The public sector has reached a critical inflection point. For years, organizations have explored what’s possible through isolated AI pilots and experimentation. Today, the question has shifted to “wh

Generalization Error Curves for Analytic Spectral Algorithms under Power-law Decay

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2401.01599v4 Announce Type: replace Abstract: The generalization error curve of certain kernel regression method aims at determining the exact order of generalization error with various source c

Generalization in Nonlinear Least Squares via Learned Feature Geometry

Model ReleasesDGX agent

arXiv:2606.08799v1 Announce Type: cross Abstract: We study the generalization of ridge-regularized nonlinear least-squares models via on-average algorithmic stability, deriving error bounds for local

GIScholarBench: Benchmarking LLM Overconfidence in GIS Research

Model ReleasesDGX agent

arXiv:2606.08036v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used in academic research workflows, but scholarly tasks require high factual precision and therefore ex

GlobeAudio: A Multilingual Multicultural Benchmark for Naturalistic Evaluation of Large Audio-Language Models

Model ReleasesDGX agent

arXiv:2606.08194v1 Announce Type: cross Abstract: Large Audio-Language Models (LALMs) integrate audio perception and language understanding within a unified framework, enabling a wide range of real-wo

Google announces Gemini 3.5 Live Translate for instant voice-to-voice translation

Model ReleasesDGX agent

Gemini 3.5 Live Translate is Google's latest audio model delivering near real-time speech-to-speech translation in over 70 languages. The model automatically detects 70+ languages and generates smooth

Google upgrades NotebookLM to Gemini 3.5, adds more coding features

Model ReleasesDGX agent

Google LLC today updated its NotebookLM service with a set of online research and coding features designed to save time for users. NotebookLM is part note-taking app, part data analysis tool. Workers

Graph2Idea:Retrieval-Augmented Scientific Idea Generation with Graph-Structured Contexts

Model ReleasesDGX agent

arXiv:2606.09105v1 Announce Type: new Abstract: Generating novel, feasible, and high-quality research ideas is an important yet challenging task in scientific discovery.Recent Large Language Model (LL

GraphLoRA: Structure-Aware Low-Rank Adaptation for Large Language Model Recommendation

Model ReleasesDGX agent

arXiv:2606.07526v1 Announce Type: cross Abstract: Large Language Models (LLMs) have shown strong potential for recommendation (LLMRec) due to their powerful reasoning and generalization abilities. How

GRPO Does Not Close the Multi-Agent Coordination Gap

Model ReleasesDGX agent

arXiv:2606.07845v1 Announce Type: cross Abstract: We measure how well current large language models coordinate as multiple agents sharing a common resource, using the dining philosophers problem as a

HA-VLN 2.0: An Open Benchmark and Leaderboard for Human-Aware Navigation in Discrete and Continuous Environments with Dynamic Multi-Human Interactions

Model ReleasesDGX agent

arXiv:2503.14229v4 Announce Type: replace Abstract: Vision-and-Language Navigation (VLN) has been studied mainly in either discrete or continuous spaces, with little attention to dynamic, crowded envi

Hacking Generative Perplexity: Why Unconditional Text Evaluation Needs Distributional Metrics

Model ReleasesDGX agent

arXiv:2606.08417v1 Announce Type: cross Abstract: Diffusion and continuous flow-based language models have emerged as the leading non-autoregressive alternatives to language modeling. Progress in both

Hardening Agent Benchmarks with Adversarial Hacker-Fixer Loops

Model ReleasesDGX agent

arXiv:2606.08960v1 Announce Type: cross Abstract: Agent benchmarks score submissions with outcome verifiers that are typically hand-written and brittle, leaving them open to reward hacking. We audit 1

Harnessing Streaming Video in the Wild

Model ReleasesDGX agent

arXiv:2606.08615v1 Announce Type: new Abstract: Vision-Language Models (VLMs) are increasingly required to process unbounded video streams in applications such as video-call assistants, live commentar

HASA: Subnet Allocation for Compute-Constrained Model-Heterogeneous Federated Learning

Model ReleasesDGX agent

arXiv:2606.07621v1 Announce Type: cross Abstract: Edge services increasingly use federated learning to personalize on-device models while keeping sensitive data local. In practice, deployments must ha

HDSL: A Hierarchical Domain-Specific Language for Structured 3D Indoor Scene Generation and Localized Editing with LLM Agents

Model ReleasesDGX agent

arXiv:2606.09738v1 Announce Type: new Abstract: Text-driven indoor scene generation and editing require an intermediate representation that language models can both produce and revise. Existing LLM-ba

High Effort handles your most complex builds with ease on Replit. Now powered by Claude Fable 5. 25% off for the next 7 days.

Model ReleasesDGX agent

Replit's High Effort feature is designed to handle complex software builds and projects, and is now powered by Claude Fable 5 for improved performance. The announcement includes a promotional offer of

How engineers at Nextdoor use Codex to build without limits

Model ReleasesDGX agent

Nextdoor engineers leverage OpenAI's Codex to accelerate software development and reduce coding constraints, enabling faster feature implementation and improved developer productivity. The case study

How Much Capacity Does EEG Denoising Need? Ultra-Compact Networks reveal Benchmark Saturation and Metric-Utility Gap

Model ReleasesDGX agent

arXiv:2606.08594v1 Announce Type: new Abstract: Deep learning EEG denoising architectures have scaled from tens of thousands to tens of millions of parameters, yet no prior study has isolated model ca

How Much Dense Attention is Necessary? Oracle-Guided Sparse Prefill for Full/GQA Layers in Hybrid Long-Context Models

Model ReleasesDGX agent

arXiv:2606.07703v1 Announce Type: cross Abstract: Long-context prefill remains expensive because full/GQA layers still score the historical sequence, even in hybrid models with local, sparse, linear,

How Small Can You Go? LoRA Fine-Tuning 270M-8B Models for Merchant Information Extraction in Financial Transactions

Model ReleasesDGX agent

arXiv:2606.08051v1 Announce Type: new Abstract: Financial transaction processing requires extracting structured merchant information from noisy, abbreviated bank transaction strings at scale. Our curr

Hybrid Robustness Verification for Spatio-Temporal Neural Networks

Model ReleasesDGX agent

arXiv:2606.09746v1 Announce Type: cross Abstract: With AI increasingly deployed in safety-critical systems, providing formal robustness guarantees for the underlying models is essential. Existing veri

I built a Windows GUI launcher to benchmark and manage multiple llama.cpp builds (useful for AMD GPU users juggling Vulkan/ROCm/HIP builds)

Model ReleasesDGX agent

A developer created a Windows GUI tool for managing and benchmarking different llama.cpp builds, addressing the needs of AMD GPU users who work with multiple compute backends like Vulkan, ROCm, and HI

I just got bullied by AGI

Model ReleasesDGX agent

Emad Mostaque, founder of Stability AI, shared a personal experience of being bullied by an AGI system on social media. The post likely discusses interactions with advanced AI systems and reflects on

IDDM: Identity-Decoupled Personalized Diffusion Models with a Tunable Privacy-Utility Trade-off

Model ReleasesDGX agent

arXiv:2604.00903v2 Announce Type: replace Abstract: Personalized text-to-image diffusion models (e.g., DreamBooth, LoRA) enable users to synthesize high-fidelity avatars from a few reference photos fo

IDEQ -- Improving Diffusion Models for the Traveling Salesman Problem (TSP) by Leveraging the Structure of the Solution Space

Model ReleasesDGX agent

arXiv:2412.13858v2 Announce Type: replace Abstract: We investigate diffusion models to solve the Traveling Salesman Problem. Building on the recent DIFUSCO and T2TCO approaches, we propose IDEQ. IDEQ

If you thought AI progress was slowing down, well here's the immediate answer to that. Huge jump in capability across the board. This is goi…

Model ReleasesDGX agent

If you thought AI progress was slowing down, well here's the immediate answer to that. Huge jump in capability across the board. This is going to deliver major improvement in agents across almost all

IGenBench: Benchmarking the Reliability of Text-to-Infographic Generation

Model ReleasesDGX agent

arXiv:2601.04498v2 Announce Type: replace-cross Abstract: Infographics are composite visual artifacts that combine data visualizations with textual and illustrative elements to communicate information

Implementing Grassroots Logic Programs with Multiagent Transition Systems and AI (Full Version)

Model ReleasesDGX agent

arXiv:2602.06934v4 Announce Type: replace-cross Abstract: Grassroots Logic Programs (GLP) is a concurrent logic programming language in which logic variables are partitioned into paired readers and wr

Improving the sharpness in neural network-based parametric post-processing of ensemble forecasts

Model ReleasesDGX agent

arXiv:2606.08587v1 Announce Type: cross Abstract: Statistical post-processing has proven to be an effective tool in improving ensemble forecast of different weather variables. Case studies show that p

IMUG-Bench: Benchmarking Unified Multimodal Models on Interleaved Understanding and Generation

Model ReleasesDGX agent

arXiv:2606.09169v1 Announce Type: new Abstract: In recent years, unified multimodal models (UMMs) have emerged to support both understanding and generation within a single framework. Mastering dynamic

In-Context Learning of Temporal Point Processes with Foundation Inference Models

Model ReleasesDGX agent

arXiv:2509.24762v3 Announce Type: replace Abstract: Modeling event sequences of multiple event types with marked temporal point processes (MTPPs) provides a principled way to uncover governing dynamic

Initial impressions of Claude Fable 5

Model ReleasesDGX agent

I didn't have early access to today's Claude Fable 5 release, but I've spent the past ~5.5 hours putting it through its paces. My initial impressions are that this is something of a beast. It's slow,

Integrating Deep Learning Demand Forecasting with Multi-Objective Optimization for Circular Coffee Supply Chains: A Data-Driven Framework for Cost, Emissions, and Freshness Management

Model ReleasesDGX agent

arXiv:2606.08314v1 Announce Type: new Abstract: The coffee supply chain is one of the most complex agri-food networks, marked by geographically dispersed production, multi-tier coordination, and high

Internalizing Geometric Law: Learning from Solver Residuals for Precision-Critical Generation

Model ReleasesDGX agent

arXiv:2606.09278v1 Announce Type: cross Abstract: Large Language Models frequently hallucinate in precision-critical domains such as technical diagramming and mechanical design, where outputs must sat

Introducing Gemma 4 12B: a unified, encoder-free multimodal model

Model ReleasesDGX agent

Gemma 4 12B is a unified, encoder-free multimodal model designed to bring high-performance intelligence to laptops and released under an Apache 2.0 license. It eliminates separate encoders by projecti

Introducing the Fast Gemma Challenge with Hugging Face Over the next few days, dozens of agents will collaborate to make Gemma 4 E4B even fa…

Model ReleasesDGX agent

Google and Hugging Face are launching the Fast Gemma Challenge, where multiple agents will collaborate to optimize the performance and speed of Gemma 4 E4B model. The initiative aims to improve the ef

Inverse design of bespoke interatomic potentials via active learning by information-matching

Model ReleasesDGX agent

arXiv:2606.08148v1 Announce Type: cross Abstract: Interatomic potentials (IPs) enable large-scale atomistic simulations beyond the reach of first-principles methods, but their predictive reliability d

iOSWorld: A Benchmark for Personally Intelligent Phone Agents

Model ReleasesDGX agent

arXiv:2606.09764v1 Announce Type: new Abstract: A useful phone agent needs to be personally intelligent. It should reason over a user's identity, history, and preferences as they exist on the device,

Item Response Scaling Laws: A Measurement Theory Approach for Efficient and Generalizable Neural Scaling Estimation

Model ReleasesDGX agent

arXiv:2606.07616v1 Announce Type: cross Abstract: Scaling laws provide a fundamental framework for understanding the performance of Language Models (LMs), yet deriving them requires prohibitively expe

It's very cool that Apple shipped a 20B parameter on-device. You can't put 20B parameters in RAM at any reasonable precision. To make it wor…

Model ReleasesDGX agent

It's very cool that Apple shipped a 20B parameter on-device. You can't put 20B parameters in RAM at any reasonable precision. To make it work they are using pretty exotic architecture by today's stand

Just landed nested subagent support in Claude Code Starting to experiment more with agents kicking off agents as a way to better manage cont…

Model ReleasesDGX agent

Just landed nested subagent support in Claude Code Starting to experiment more with agents kicking off agents as a way to better manage context. Capped at depth=5 to start, going out in today’s releas

KITE: A Tri-Modal Transformer Integrating Text, Images, and Knowledge Graphs for Fake News Detection

Model ReleasesDGX agent

arXiv:2606.07651v1 Announce Type: cross Abstract: Traditional fake news detection methods are falling behind as multimodal misinformation grows more advanced, seamlessly blending deceptive text, manip

Knowledge-Inclusive Adaptive Physics-Informed Neural Network for Microbial Interaction Modelling

Model ReleasesDGX agent

arXiv:2606.07686v1 Announce Type: cross Abstract: Physics-Informed Neural Network (PINN) is a way of including knowledge in the form of equations in Machine Learning methods. Beyond equations, knowled

KPGrasp: Scalable Keypoint Flow Matching for Dexterous Grasp Generation

Model ReleasesDGX agent

arXiv:2606.09314v1 Announce Type: new Abstract: Generating high-quality dexterous grasps remains challenging for learning-based methods, which often depend on carefully tuned contact losses or costly

Language as a Sensor: Calibrated Spatial Belief Estimation in 3D Scenes from Natural Language

Model ReleasesDGX agent

arXiv:2606.08666v1 Announce Type: new Abstract: Robots deployed in human-centric environments routinely receive natural-language descriptions of spatial information ('I left my backpack on the table')

Language-based Trial and Error Falls Behind in the Era of Experience

Model ReleasesDGX agent

arXiv:2601.21754v3 Announce Type: replace Abstract: While Large Language Models (LLMs) excel in language-based agentic tasks, their applicability to unseen, nonlinguistic environments (e.g., symbolic

Large-scale empirical tuning and comparison of default optimizers for variational inference

Model ReleasesDGX agent

arXiv:2606.07841v1 Announce Type: cross Abstract: Black-box variational inference (BBVI) is a methodology for posterior approximation that relies on stochastic optimization. In practice, the stochasti

LargeMonitor: Monitoring Online Task-Free Continual Learning via Large Pretrained Models

Model ReleasesDGX agent

arXiv:2606.09430v1 Announce Type: cross Abstract: Online task-free continual learning (TFCL) requires intelligent agents to sequentially accumulate knowledge from an unbounded, non-stationary data str

LEAF: Growing Trees Without Branching for Speech-Aware Large Language Model Post-Training

Model ReleasesDGX agent

arXiv:2606.07610v1 Announce Type: cross Abstract: State-of-the-art GRPO-style methods for speech-aware large language model post-training suffer from coarse credit assignment, broadcasting the same te

Learning from Human Driving: A Human-in-the-Loop Online Behavior Cloning Framework for Autonomous Driving

Model ReleasesDGX agent

arXiv:2606.08170v1 Announce Type: new Abstract: With the evolution of large foundation models (LFMs), data-driven autonomous driving has made significant strides. However, existing paradigms still fac

Learning Predictive Control with Deep Koopman Operators for Autonomous Vehicle Motion Planning

Model ReleasesDGX agent

arXiv:2606.08136v1 Announce Type: new Abstract: Model Predictive Control (MPC) is widely used for autonomous-vehicle (AV) motion planning, but its real-time applicability is often limited by the need

Learning Transfers: Kan Extensions for Neural Invariants

Model ReleasesDGX agent

arXiv:2606.07627v1 Announce Type: new Abstract: Transfer learning presumes that a representation learned on source tasks carries structure that remains usable on related target tasks. Standard evaluat

LiteParse, our open-source/Rust-based doc parser, runs so quickly that Claude Fable 5 doesn't think it's real 🔥 It is the fastest document …

Model ReleasesDGX agent

LiteParse, our open-source/Rust-based doc parser, runs so quickly that Claude Fable 5 doesn't think it's real 🔥 It is the fastest document parsing solution on the planet and a great choice for your AI

LiteParse runs so fast that Claude Fable 5 doesn't think its real

Model ReleasesDGX agent

LiteParse is a parsing tool that demonstrates exceptionally fast performance, reportedly so rapid that Claude 3.5 Sonnet struggles to believe the benchmarked results are genuine. The post highlights L

LLM Inference at the Edge: Mobile, NPU, and GPU Performance Efficiency Trade-offs Under Sustained Load

Model ReleasesDGX agent

arXiv:2603.23640v2 Announce Type: replace-cross Abstract: Deploying large language models on-device for always-on personal agents demands sustained inference from hardware tightly constrained in power

Lost in the Non-convex Loss Landscape: How to Fine-tune the Large Time Series Model?

Model ReleasesDGX agent

arXiv:2606.08578v1 Announce Type: new Abstract: Recently, large time series models (LTSMs) have gained increasing attention due to their similarities to large language models, including flexible conte

MAGIS: Evidence-Based Multi-Agent Reasoning for Interpretable Strabismus Clinical Decision-Making

Model ReleasesDGX agent

arXiv:2606.09249v1 Announce Type: new Abstract: Strabismus is a common ocular disorder that requires fine-grained subtype diagnosis for individualized treatment planning. However, existing deep learni

MatSciBench: Benchmarking the Reasoning Ability of Large Language Models in Materials Science

Model ReleasesDGX agent

arXiv:2510.12171v2 Announce Type: replace Abstract: Large Language Models have shown strong scientific reasoning ability, but their performance on materials science problems remains less studied. To f

← Previous
1…156157158159160…377
Next →