AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

84,630Total entries
1Added by human
84,629Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,575 results
9 Jul 2026

Aurora 1.5: Extending open foundation models for weather and Earth-system applications

ApplicationsDGX agent

Aurora 1.5 adds 22 more variables, hourly temporal resolution, and probabilistic ensemble forecasting to the Aurora foundation model, making it more useful for real-world weather, climate, and energy

DYNA-PRUNER: Input-Adaptive Data-Model Co-Pruning for Efficient and Scalable Spatio-Temporal Media Prediction

HardwareDGX agent

arXiv:2606.15346v2 Announce Type: replace Abstract: Spatio-temporal prediction supports radar/satellite nowcasting and city-scale traffic monitoring, but modern models are often too expensive for real

ECHO: Ego-Centric modeling of Human-Object interactions

ResearchDGX agent

arXiv:2508.21556v3 Announce Type: replace Abstract: Modeling human-object interactions (HOI) from an egocentric perspective is a critical yet challenging task, particularly when relying on sparse sign

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

EdgeCompress: Coupling Multidimensional Model Compression and Dynamic Inference for EdgeAI

ResearchDGX agent

arXiv:2607.06982v1 Announce Type: new Abstract: Convolutional neural networks (CNNs) have demonstrated encouraging results in image classification tasks. However, the prohibitive computational cost of

Future Confidence Distillation in Large Language Models

AgentsDGX agent

arXiv:2607.07626v1 Announce Type: cross Abstract: Reliable confidence estimation is essential for deploying large language models (LLMs) in confidence-aware systems, where downstream decisions such as

Loving this piece: bringing delight to your daily coffee, your daily workflows, and your AI bill. @gumloop is now running open-weight models…

ApplicationsDGX agent

Loving this piece: bringing delight to your daily coffee, your daily workflows, and your AI bill. @gumloop is now running open-weight models on Fireworks in production. 80% lower cost vs. closed-lab A

Memory Scarcity, Open Models, and the Restructuring of the AI Industry, 2026-2030 -- A quantitative scenario analysis of inference economics, training-cost divergence, and infrastructure solvency

Local AiDGX agent

arXiv:2607.07207v1 Announce Type: cross Abstract: We analyze how four forces restructure the AI industry over 2026-2030: the DRAM/HBM price surge, frontier-capable open-weight models (GLM-5.2), rapid

obviously the best model we have ever produced, but also one of the best blog posts we have ever produced: https://openai.com/index/gpt-5-6/

Model ReleasesDGX agent

I cannot verify this URL or tweet as it appears to reference a future date (2075) and a non-existent OpenAI blog post about GPT-5-6. This does not match any verified OpenAI announcements or Sam Altman

Seedream 5.0 Pro is now in ComfyUI via Partner Nodes. ByteDance's latest image model brings: → Character & product consistency across edits …

Local AiDGX agent

Seedream 5.0 Pro is now in ComfyUI via Partner Nodes. ByteDance's latest image model brings: → Character & product consistency across edits → Precision region editing → Infographics, flowcharts & stru

Tensorized algorithms and scalable filtering methods for hidden Markov and factorial hidden Markov models

ApplicationsDGX agent

arXiv:2607.07008v1 Announce Type: cross Abstract: A common method for the representation and analysis of time-series data is the hidden Markov model (HMM), where each observation is associated with a

Validate the Dream Before You Trust Its Verdict: Admissibility for World-Model Simulators

SafetyDGX agent

arXiv:2607.07196v1 Announce Type: cross Abstract: Across robotics, World Models (WMs) are increasingly used to evaluate action policies by simulating the consequences of actions in an imagined world,

🤗 we just opensourced our tiny beast on @huggingface >0.9B ASR model >Apache license 2.0 >128k long-context transcription >Up to ~90-min au…

HardwareDGX agent

🤗 we just opensourced our tiny beast on @huggingface >0.9B ASR model >Apache license 2.0 >128k long-context transcription >Up to ~90-min audio input >Speaker labels + timestamps in one generation >Mul

8 Jul 2026

A Gibbs posterior sampler for inverse problem based on prior diffusion model

ResearchDGX agent

arXiv:2602.11059v2 Announce Type: replace-cross Abstract: This paper addresses the issue of inversion in cases where (1) the observation system is modeled by a linear transformation and additive error

Abductive Corroboration of Probabilistic AI Models for Forensic Synthetic Media Detection

ApplicationsDGX agent

arXiv:2607.05434v1 Announce Type: cross Abstract: Artificial Intelligence (AI) models, at their core, apply general learnings from broad datasets to individual circumstances using probabilistic behavi

Canopy: A Heterograph Foundation Model for Metabolic Engineering

TutorialsDGX agent

arXiv:2607.06224v1 Announce Type: new Abstract: Designing microbial strains that produce high-value chemicals at commercially viable titers remains a central challenge in metabolic engineering. Existi

Efficient Transfer Learning of Robot Dynamic Models Using Morphological Similarity

ResearchDGX agent

arXiv:2607.05665v1 Announce Type: new Abstract: This study proposes a neural network-based transfer learning framework for modeling the dynamics of soft, fin-actuated underwater robots. We focus on mo

Fine-Tuning Integrity for Modern Neural Networks: Structured Drift Proofs via Norm, Rank, and Sparsity Certificates

Model ReleasesDGX agent

arXiv:2604.04738v2 Announce Type: replace-cross Abstract: Fine-tuning is the dominant paradigm for adapting large machine learning models, yet current deployment pipelines provide no way to verify how

My big takeaway is that both Sol & Fable represent jumps over previous models and have opened a large gap with the next-best AIs. People wil…

ApplicationsDGX agent

My big takeaway is that both Sol & Fable represent jumps over previous models and have opened a large gap with the next-best AIs. People will have preferences for one or the other, but if you doing an

On the Redundancy of Timestep Embeddings in Diffusion Models

ResearchDGX agent

arXiv:2606.20416v2 Announce Type: replace-cross Abstract: Diffusion models rely heavily on explicit timestep embeddings to modulate the denoising process across various noise scales. In this work, we

Open the model weights, Hal! Congrats for the new round @PrimeIntellect

IndustryDGX agent

This post from Hugging Face co-founder Clem Delangue congratulates Prime Intellect on securing a new funding round and calls for them to open their model weights, likely advocating for open-source AI

PORTS: Preference-Optimized Retrievers for Tool Selection with Large Language Models

SafetyDGX agent

arXiv:2607.05441v1 Announce Type: cross Abstract: Integrating external tools with Large Language Models (LLMs) has emerged as a promising paradigm for accomplishing complex tasks. Since LLMs still str

Replication in Visual Diffusion Models: A Survey and Outlook

ApplicationsDGX agent

arXiv:2408.00001v2 Announce Type: replace-cross Abstract: Visual diffusion models have revolutionized the field of creative AI, producing high-quality and diverse content. However, they inevitably mem

Same prompt. 3 models. 6 camera moves. We put Seedance 2.0, Kling 3.0, and Grok Imagine 1.5 head-to-head on the camera techniques → Bullet t…

ApplicationsDGX agent

Same prompt. 3 models. 6 camera moves. We put Seedance 2.0, Kling 3.0, and Grok Imagine 1.5 head-to-head on the camera techniques → Bullet time → Crash zoom → Crane reveal → Dolly zoom → Object POV →

SHARC: SHAP-Based Interpretability in Machine Learning Risk Models for Regulatory Capital under ICAAP and CCAR

Local AiDGX agent

arXiv:2607.05484v1 Announce Type: cross Abstract: The adoption of non-parametric machine learning models for regulatory capital estimation introduces a fundamental governance challenge: the inability

SIEVE: Structure-Aware Data Selection for Imitation Learning with VLA Models

ResearchDGX agent

arXiv:2607.06442v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models are typically trained by imitation learning on large-scale robot demonstration datasets, but more data does not nece

Structured-Condensed Prompt Tuning in Vision-Language Models for Fine-grained Image Recognition

SafetyDGX agent

arXiv:2607.06185v1 Announce Type: new Abstract: Fine-grained image recognition poses a significant challenge due to the substantial expertise and effort required for manual annotation. Vision-language

Tuning-Free Latent Diffusion Models for Ultrahigh-Resolution Image Editing

HardwareDGX agent

arXiv:2607.06136v1 Announce Type: new Abstract: Recent diffusion-based generative models have shown impressive performance in image generation and editing. However, due to memory limitations and the h

7 Jul 2026

A Decomposable Probe for Few-Step Diffusion Models: Prompt, Latent, and Score Selectivity across Backbone Families and Distillation Paradigms

ResearchDGX agent

arXiv:2607.03256v1 Announce Type: new Abstract: Few-step distilled diffusion students cut text-to-image inference from ~50 to 1-8 network evaluations, but the quality gap is usually summarised by a si

A Random Matrix Theory Perspective on the Consistency of Diffusion Models

ResearchDGX agent

arXiv:2602.02908v2 Announce Type: replace-cross Abstract: Diffusion models trained on different, non-overlapping subsets of a dataset often produce strikingly similar outputs when given the same noise

Agentic-V2X: Small Language Model Agents for Deadline-Aware V2X Scheduling in 5G/6G Networks

Local AiDGX agent

arXiv:2607.04290v1 Announce Type: cross Abstract: Large Language Models (LLMs) are proposed as control interfaces for next-generation networks, but their latency, hallucinations, and lack of control g

Alleviating Sparse Rewards by Modeling Step-Wise and Long-Term Sampling Effects in Flow-Based GRPO

ResearchDGX agent

arXiv:2602.06422v2 Announce Type: replace Abstract: Deploying GRPO on Flow Matching models has proven effective for text-to-image generation. However, existing paradigms typically propagate an outcome

Benchmarking API Drift in LLM-Generated Quantum Code Across Successive SDK Versions

Model ReleasesDGX agent

arXiv:2607.04072v1 Announce Type: cross Abstract: Large language models can generate plausible quantum code, but it is unclear whether they can reliably target the specific software development kit (S

BVS: Bayesian Visual Search with Multimodal Large Language Model for Fine-grained Perception

ResearchDGX agent

arXiv:2607.03184v1 Announce Type: new Abstract: While Multimodal Large Language Models (MLLMs) demonstrate impressive general capabilities, they struggle with fine-grained perception in ultra-high-res

CompressedVQA-AEV: Full-Reference and No-Reference Quality Assessment Models for Asymmetric Encoded Videos

ResearchDGX agent

arXiv:2607.04606v1 Announce Type: cross Abstract: This report presents our solutions to the QoMEX 2026 Grand Challenge on Video Quality Assessment for Asymmetric Encoded Videos, comprising a full-refe

CRISP: A Spatiotemporal Camera-Radar Backbone for Driving via Forecasting-Based World-Model Pretraining

AgentsDGX agent

arXiv:2607.04541v1 Announce Type: cross Abstract: Camera-radar (CR) fusion is a practical sensing configuration for autonomous driving, but existing models are typically trained with task-specific sup

Determinants and Limits of LLM Security-Tool Orchestration: A Study with HexStrike-AI

Model ReleasesDGX agent

arXiv:2607.02873v1 Announce Type: cross Abstract: Large language model agents driving security tool suites over the Model Context Protocol are increasingly common. Yet the factors that bound their cap

Diagnosing Aerial-View Object Detectors with Foundational Image Generative Models

TutorialsDGX agent

arXiv:2607.02718v1 Announce Type: cross Abstract: Recent advances in large-scale image generative models enable photorealistic scene synthesis with controllable attributes. Beyond data augmentation, t

DistillH-Mamba: A Hypergraph-Mamba-Based Knowledge Distillation Model for Efficient Impact Fall Detection

ResearchDGX agent

arXiv:2607.03156v1 Announce Type: new Abstract: Falls among the elderly represent a significant public health concern due to their prevalence, consequences, and societal burden. While deep learning ha

Do ECG Foundation Models Transfer to Rare Cardiac Diseases? Evidence from Brugada Syndrome Detection

SafetyDGX agent

arXiv:2607.03009v1 Announce Type: new Abstract: Background: Foundation models (FMs) trained on large-scale unlabeled physiological data have emerged as a promising paradigm for medical artificial inte

Don't Commit Alone: Joint Token Commitment in Diffusion Large Language Models

ResearchDGX agent

arXiv:2607.04469v1 Announce Type: new Abstract: Diffusion large language models (dLLMs) commit multiple tokens per denoising step by decoding each selected position independently from the shared conte

DuplexChat: Constructing Speaker-Separated Full-Duplex Dialogue Speech at Scale for Spoken Dialogue Language Modeling

ResearchDGX agent

arXiv:2607.04941v1 Announce Type: new Abstract: Full-duplex spoken dialogue models are trained on conversational speech in which each speaker is represented as a separate stream, but existing large-sc

Efficient bias mitigation in T2I diffusion models using Concept Graphs

SafetyDGX agent

arXiv:2607.03397v1 Announce Type: new Abstract: Text-to-Image diffusion models often propagate harmful bias inherited from the training data. Existing bias mitigation techniques typically intervene on

European banking watchdogs ECB and ESRB warn that frontier AI models pose 'systemic risks to the financial system', giving lenders four months to prepare (Financial Times)

IndustryDGX agent

Financial Times: European banking watchdogs ECB and ESRB warn that frontier AI models pose “systemic risks to the financial system”, giving lenders four months to prepare — IT weaknesses could be expl

Forethought: Verifiable Reasoning from Neurosymbolic Primitive Programming

Model ReleasesDGX agent

arXiv:2607.04096v1 Announce Type: new Abstract: Current agentic workflows usually involve decomposing user requests into sequences of tool calls with correctly resolved parameters, the results of whic

From Global to Local: Efficient Regional Weather Downscaling with Global Weather Foundation Model

Local AiDGX agent

arXiv:2607.03279v1 Announce Type: new Abstract: Accurate regional weather prediction requires resolving fine-scale structure while remaining consistent with global dynamics. Traditional limited area m

HALO-WA: Hybrid-Attention Latent-Guided Online Reinforcement Learning for World-Action Models

SafetyDGX agent

arXiv:2607.04265v1 Announce Type: cross Abstract: World-action (WA) models can generate long-horizon action chunks for general-purpose robotic manipulation, but they remain vulnerable to calibration,

Handwriting Trajectory Recovery with Diffusion Models

ResearchDGX agent

arXiv:2607.03422v1 Announce Type: new Abstract: Recovering online pen trajectories from offline handwriting images, often referred to as handwriting trajectory recovery (stroke recovery), is an offlin

Hierarchical Anti-Aesthetics: Protecting Facial Privacy against Customized Diffusion Models

TutorialsDGX agent

arXiv:2607.02038v2 Announce Type: replace Abstract: The rise of customized diffusion models has fueled a boom in personalized visual content creation, but it also introduces serious risks of malicious

HiSAC: Hierarchical Sparse Activation Compression for Ultra-long Sequence Modeling in Recommenders

ApplicationsDGX agent

arXiv:2602.21009v2 Announce Type: replace-cross Abstract: Modern recommender systems leverage ultra-long user behavior sequences to capture dynamic preferences, but end-to-end modeling is infeasible i

How Much of the Routing Gap Is Real? Decomposing the Router-to-Oracle Gap into Reproducible Specialist Advantage and Single-Draw Label Noise

Model ReleasesDGX agent

arXiv:2607.03436v1 Announce Type: new Abstract: Routing among large language models (LLMs) promises better quality at lower cost, motivated by the reported gap between learned routers and a per-instan

Human-like Object Grouping in Self-supervised Vision Transformers

Model ReleasesDGX agent

arXiv:2603.13994v2 Announce Type: replace-cross Abstract: Vision foundation models trained with self-supervised objectives achieve strong performance across diverse tasks and exhibit emergent object s

Identifiability Without Gaussianity: Symbolic World Models and Near-Infinite Temporal Consistency

SafetyDGX agent

arXiv:2606.12471v2 Announce Type: replace-cross Abstract: Klindt, LeCun, and Balestriero (arXiv:2605.26379) proved that Joint-Embedding Predictive Architectures (JEPAs) achieve linear identifiability,

InverseScope: Scalable Activation Inversion for Interpreting Large Language Models

ResearchDGX agent

arXiv:2506.07406v3 Announce Type: replace-cross Abstract: Understanding the internal representations of large language models (LLMs) is a central challenge in interpretability research. Existing featu

@__JohnNguyen__ and I are presenting the Beyond Language Modeling paper as an ICML spotlight in 30 minutes at the 10:30 AM poster session! I…

ResearchDGX agent

@__JohnNguyen__ and I are presenting the Beyond Language Modeling paper as an ICML spotlight in 30 minutes at the 10:30 AM poster session! I’ll also be at the AMI Mixer on Thursday and hanging around

Joint distribution of upstream runoff governs downstream river-discharge prediction uncertainty in distributed ML models

ApplicationsDGX agent

arXiv:2607.03217v1 Announce Type: new Abstract: Uncertainty quantification of hydrological predictions is necessary to inform operational decisions. Recent generative machine-learning methods have adv

Learning Taxonomic Trees with Hierarchical Representation Regularization for Large Multimodal Models

ResearchDGX agent

arXiv:2607.02909v1 Announce Type: cross Abstract: Taxonomies provide key information about the semantic relationships between concepts and the inherent organization of vision and language. Despite the

NRT-Bench: Benchmarking Multi-Turn Red-Teaming of LLM Operator Agents in Safety-Critical Control Rooms

Model ReleasesDGX agent

arXiv:2606.20408v3 Announce Type: replace-cross Abstract: Large language model (LLM) agents are increasingly proposed as supervisory components for safety-critical systems, yet their robustness under

NVIDIA and Hugging Face Bring New Models and Frameworks to LeRobot for the Open Robotics Community

HardwareDGX agent

Open source AI has shown how quickly developers can innovate when models, data and tools are shared. Robotics has the same opportunity, but advancements in physical AI development can still be gated b

OpenRouter: Chinese AI models have drawn 30%+ of token use by US companies each week since February 8, peaking at 46%, up from 11% over the previous 12 months (Kai Nicol-Schwarz/CNBC)

IndustryDGX agent

Kai Nicol-Schwarz / CNBC: OpenRouter: Chinese AI models have drawn 30%+ of token use by US companies each week since February 8, peaking at 46%, up from 11% over the previous 12 months — Chinese-built

PLoRA: Efficient Concurrent LoRA Training for Large Language Models

ResearchDGX agent

arXiv:2508.02932v2 Announce Type: replace Abstract: Low-Rank Adaptation (LoRA) has gained popularity as a fine-tuning approach for Large Language Models (LLMs) due to its low resource requirements and

← Previous
1…173174175176177…1010
Next →