AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,115
  • Agents7,313
  • Applications5,228
  • Concepts5
  • Hardware1,762
  • Industry6,105
  • Local Ai4,756
  • Model Releases22,759
  • Research19,333
  • Safety12,889
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,115
  • Agents7,313
  • Applications5,228
  • Concepts5
  • Hardware1,762
  • Industry6,105
  • Local Ai4,756
  • Model Releases22,759
  • Research19,333
  • Safety12,889
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

85,115Total entries
1Added by human
85,114Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,986 results
23 Jun 2026

Demystifying Training-Time Augmentation for Data-Constrained Language Model Pretraining

ResearchDGX agent

arXiv:2606.16246v2 Announce Type: replace Abstract: As AI labs approach a data ceiling where compute capacity outpaces the rate of new high-quality text generation, language model pretraining is shift

Efficient Network Inference via Hardware-Aware Architecture Search, Model Pruning & Quantization

ResearchDGX agent

arXiv:2606.23210v1 Announce Type: new Abstract: Embedded global navigation satellite system (GNSS) interference monitoring requires fast and memory-efficient inference to process large volumes of raw

EnTrust: Modeling Inter-Modal Conflict for Trustworthy Multimodal Medical Image Analysis

Local AiDGX agent

arXiv:2606.21384v1 Announce Type: new Abstract: Multimodal medical imaging fuses complementary anatomical and functional information, yet modalities frequently disagree in pathologically heterogeneous

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

FlowTrain: Flow-Based Decoupled Training for Industrial-Grade Vision-Language Models

ApplicationsDGX agent

arXiv:2606.23087v1 Announce Type: new Abstract: Industrial-grade distributed training of vision-language models (VLMs) remains far less efficient than that of unimodal LLMs. Existing solutions either

Foresight: Failure Detection for Long-Horizon Robotic Manipulation with Action-Conditioned World Model Latents

ApplicationsDGX agent

arXiv:2606.23085v1 Announce Type: new Abstract: Long-horizon tasks are common in real-world robotic deployments, yet failure detection for such tasks remains underexplored. Detecting failures in long-

Foundation Models for Epileptogenic Zone Identification in Drug-Resistant Epilepsy

ResearchDGX agent

arXiv:2606.22657v1 Announce Type: new Abstract: Accurate identification of the epileptogenic zone (EZ) is essential for seizure freedom after resective surgery in drug-resistant epilepsy, yet seizure

From Pixels to Concepts: Growing Rich 3D Semantic Scene Graph Forests utilizing Foundation Models

ApplicationsDGX agent

arXiv:2606.23312v1 Announce Type: new Abstract: Operating in complex real-world environments requires robots to understand their surroundings on a functional semantic level. This demands a detailed mu

GreenRFM: Learning a resource-efficient radiology vision-language foundation model via supervision-centric pre-training

Local AiDGX agent

arXiv:2603.06467v2 Announce Type: replace Abstract: Radiology foundation models (RFMs) have largely inherited the scale-first recipe of natural-image vision--language pre-training. This recipe is diff

Krea 2 is back and this time the weights are OPEN. This open source release ships with two models designed to work together — @krea_ai 2 RAW…

Local AiDGX agent

Krea 2 is back and this time the weights are OPEN. This open source release ships with two models designed to work together — @krea_ai 2 RAW and Krea 2 Turbo. RAW is for training, Turbo is for inferen

Learning with Multiple Correct Answers -- Regret Bounds under Different Feedback Models

ResearchDGX agent

arXiv:2602.09402v2 Announce Type: replace Abstract: We study the problem of learning with multiple correct answers, where each instance admits a set of valid labels. We primarily focus on the online s

Meta copied slot machines to addict kids to Instagram. Now Zuckerberg is turning his company into a prediction market. Meta’s business model…

SafetyDGX agent

Meta copied slot machines to addict kids to Instagram. Now Zuckerberg is turning his company into a prediction market. Meta’s business model is profiting from addiction—kids, gamblers, & more. Stop it

Mind the Noise: Sensitivity of Transformer-based Interaction-Aware Trajectory Prediction Models to Noisy Data

Local AiDGX agent

arXiv:2606.21344v1 Announce Type: cross Abstract: Trajectory prediction allows autonomous vehicles to anticipate the future behavior of surrounding objects (or agents) and, accordingly, maximize the s

Multi-Year-to-Decadal Temperature Prediction using a Machine Learning Model-Analog Framework

SafetyDGX agent

arXiv:2502.17583v2 Announce Type: replace-cross Abstract: Multi-year-to-decadal climate predictions are a key tool in understanding the range of potential regional climate futures. Here, we present a

Numerical stability analysis of large language models

Local AiDGX agent

arXiv:2503.10251v2 Announce Type: replace-cross Abstract: Transformers are the state-of-the-art architecture for large language models, and a key to their scalability is the strategic usage of low-pre

Online World Modeling Enables Real-World Inverse Reinforcement Learning from Observation

TutorialsDGX agent

arXiv:2602.24121v2 Announce Type: replace Abstract: Current methods in robot learning are fundamentally bottlenecked by one or more of: hand-designed rewards, simulation modeling, or action supervisio

Physics-Guided Spatiotemporal State Space Modeling for Lookahead Molten Pool Segmentation in Laser Wire-Feed Welding

ResearchDGX agent

arXiv:2606.23028v1 Announce Type: new Abstract: Real-time weld-pool perception is critical for closed-loop control in laser wire-feed welding, where sensing, computation, and actuator response introdu

Pre-Generation Hallucination Detection in Large Language Models via Soft-Target Attention Probing

ResearchDGX agent

arXiv:2606.21917v1 Announce Type: cross Abstract: Detecting hallucination risk before generation enables abstention, retrieval augmentation, and routing decisions without incurring the cost of decodin

RECALL: Recovery Experience Collection for Active Lifelong Learning in Vision-Language-Action Models

SafetyDGX agent

arXiv:2606.23617v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models are commonly fine-tuned through passive imitation learning, where additional demonstrations are collected for task

Recency/Frequency Adaptive KV Caching for Large Language Model Serving

SafetyDGX agent

arXiv:2606.21238v1 Announce Type: cross Abstract: Key-value (KV) caching is a powerful technique for accelerating large language model inference and generation. Inference workloads are large and diver

Rethinking the Adaptation of Vision Foundation Models for Efficient Cell Segmentation

ResearchDGX agent

arXiv:2606.21913v1 Announce Type: new Abstract: Cell segmentation is critical for computational pathology and biomedical discovery. While recent Vision Foundation Models (VFMs) have demonstrated remar

Robust Diffusion Models via Divergence-Induced Weighted Denoising

ResearchDGX agent

arXiv:2606.22521v1 Announce Type: cross Abstract: We show that replacing the standard MSE denoising loss in diffusion models with a nonlinear transformation induced by an f-divergence yields a simple

Semi-Supervised Vision-Language-Action Model

SafetyDGX agent

arXiv:2606.21493v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models enable robots to predict actions directly from visual observations and language instructions, but adapting them to n

SignVLA: Real-Time Sign Language-Guided Robotic Manipulation via Attention LSTM and Vision-Language-Action Models

SafetyDGX agent

arXiv:2606.20857v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models enable robots to execute manipulation tasks from natural-language instructions grounded in visual observations. Ho

T-VSS: Test-Time Visual Subspace Steering for Adversarial Robustness of Vision-Language Models

ResearchDGX agent

arXiv:2606.23132v1 Announce Type: new Abstract: Vision-language models (VLMs) achieve strong zero-shot recognition, but they remain highly vulnerable to adversarial perturbations. Recent test-time ada

The Trilemma of Truth in Large Language Models

ResearchDGX agent

arXiv:2506.23921v5 Announce Type: replace-cross Abstract: The public often attributes human-like qualities to large language models (LLMs), assuming that they 'know' certain things. In reality, LLMs e

The upcoming @Seedance 2.5 model looking insane as well with multi asset input, way longer outputs etc I would expect @grok imagine to keep …

IndustryDGX agent

The upcoming @Seedance 2.5 model looking insane as well with multi asset input, way longer outputs etc I would expect @grok imagine to keep pace and real time of this quality by end of next year (!) E

today, we release the open weights of Krea 2. welcome Krea 2 Raw and Krea 2 Turbo, an undistilled model from mid-training meant to be fine-t…

IndustryDGX agent

today, we release the open weights of Krea 2. welcome Krea 2 Raw and Krea 2 Turbo, an undistilled model from mid-training meant to be fine-tuned, and a fast distilled version with a wide aesthetic div

TraceMark-LDM: Authenticatable Watermarking for Latent Diffusion Models via Binary-Guided Rearrangement

ResearchDGX agent

arXiv:2503.23332v2 Announce Type: replace Abstract: Image generation algorithms are increasingly integral to diverse aspects of human society, driven by their practical applications. However, insuffic

Variance-Tilted Diffusion Models for Diverse Sampling

ResearchDGX agent

arXiv:2606.22239v1 Announce Type: cross Abstract: Diffusion models are typically sampled independently, even when the downstream objective is to obtain a diverse set of candidates. We introduce a vari

VegSim: A Geospatial World Model for Scenario-Conditioned Vegetation Simulation

ApplicationsDGX agent

arXiv:2606.21961v1 Announce Type: new Abstract: Vegetation monitoring under climate stress requires answering not only how it will evolve given the expected weather, but how it would respond to altern

VLA-FAIL: Efficient Task Failure Detection for Finetuned Vision-Language-Action Models

ApplicationsDGX agent

arXiv:2606.21386v1 Announce Type: cross Abstract: Vision-language-action models (VLAs) achieve state-of-the-art performance on many robotic manipulation tasks, yet they can still behave unpredictably

What Does a Chemical Language Model Know About Molecules?

TutorialsDGX agent

arXiv:2606.23443v1 Announce Type: new Abstract: Chemical language models (cLMs) are widely assumed to learn surface-level syntactic patterns rather than learning meaningful molecular semantics. Here,

What if? Emulative Simulation with World Models for Situated Reasoning

SafetyDGX agent

arXiv:2603.06445v2 Announce Type: replace Abstract: Situated reasoning often relies on active exploration, yet in many real-world scenarios such exploration is infeasible due to physical constraints o

22 Jun 2026

Sakana Fugu Ultra is live on AI Gateway. Mythos-class intelligence in a single call, with a whole pool of models behind it. 𝚖𝚘𝚍𝚎𝚕: '𝚜…

ResearchDGX agent

Sakana Fugu Ultra is live on AI Gateway. Mythos-class intelligence in a single call, with a whole pool of models behind it. 𝚖𝚘𝚍𝚎𝚕: '𝚜𝚊𝚔𝚊𝚗𝚊/𝚏𝚞𝚐𝚞-𝚞𝚕𝚝𝚛𝚊' https://vercel.com/changelog/sakana-fugu-ultra-no

To be clear: 1. No it can't. I've used Fable while it was available, it was a good model but still less than 1% of the way there. 2. If it c…

ResearchDGX agent

To be clear: 1. No it can't. I've used Fable while it was available, it was a good model but still less than 1% of the way there. 2. If it could, that fact would generally benefit SaaS companies, not

Valve says Steam Machine, its new living room-friendly PC, will start at $1,049 for the 512GB base model, and go on sale starting June 29 (Jay Peters/The Verge)

IndustryDGX agent

Jay Peters / The Verge: Valve says Steam Machine, its new living room-friendly PC, will start at $1,049 for the 512GB base model, and go on sale starting June 29 — You can register your interest start

21 Jun 2026

A year ago this would have been an obvious closed-model task. Now GLM-5.2 can read the issue, reason through the scene, patch the code, and …

ToolsDGX agent

A year ago this would have been an obvious closed-model task. Now GLM-5.2 can read the issue, reason through the scene, patch the code, and keep moving on Together AI. @togethercompute + @Zai_org GLM

If you’ve been thinking about training models or like the idea but don’t know where to start This is one of the best reads worth your time, …

IndustryDGX agent

If you’ve been thinking about training models or like the idea but don’t know where to start This is one of the best reads worth your time, it’s like being able to go back in time and read your own no

20 Jun 2026

If AI self-improvement, even in a very limited way, is possible, the cadence of shipping both AI products/harnesses & models should go up. T…

ApplicationsDGX agent

If AI self-improvement, even in a very limited way, is possible, the cadence of shipping both AI products/harnesses & models should go up. This appears to be happening at Anthropic & OpenAI, but not f

There will be an open source fable-level model that runs on a base MacBook mini / Air or equivalent. I don’t think people have realised this…

IndustryDGX agent

Emad Mostaque predicts that open-source AI models will soon reach a capability level ('fable-level') where they can run efficiently on consumer-grade MacBook Air/Mini computers without specialized har

Traded my Range Rover for a @Tesla Model Y and picked it up two weeks ago. Turned on FSD today after the 14.3.3 update. If you don’t believe…

IndustryDGX agent

Traded my Range Rover for a @Tesla Model Y and picked it up two weeks ago. Turned on FSD today after the 14.3.3 update. If you don’t believe in magic, I don’t know what to tell you. I’ve always been f

We, along with @zhijianliu_ and @jianchen1799, are excited to share these models on @huggingface for the entire community to use. We look fo…

IndustryDGX agent

We, along with @zhijianliu_ and @jianchen1799, are excited to share these models on @huggingface for the entire community to use. We look forward to getting speculative decoding speedups into the hand

19 Jun 2026

Clients of SemiAnalysis Memory and Accelerator model knew about Rubin Ultra 16 Hi to 12 Hi cut in March :)

HardwareDGX agent

Clients of SemiAnalysis Memory and Accelerator model knew about Rubin Ultra 16 Hi to 12 Hi cut in March :) LMAO Rubin Ultra's HBM literally got downgraded to 12 Hi, and they are not even using HB yet

11 Jun 2026

ConsistencyPlanner: Real-time Planning with Fast-Sampling Consistency Models

SafetyDGX agent

arXiv:2606.11569v1 Announce Type: cross Abstract: Closed-loop planning in complex, real-world driving scenarios presents a critical challenge for autonomous driving systems. While traditional rule-bas

DAM-VLA: Decoupled Asynchronous Multimodal Vision Language Action model

ApplicationsDGX agent

arXiv:2606.12105v1 Announce Type: cross Abstract: Vision-language-action (VLA) models inherit a shared synchronous clock from vision-language pretraining, processing every input at one rate. This is m

From Architecture to Output: Structural Origins of Hallucination in Large Language Models and the Amplifying Role of Data

SafetyDGX agent

arXiv:2606.07537v1 Announce Type: cross Abstract: Large language models hallucinate--producing fluent, confident, factually wrong outputs--with a consistency that persists across generations and scale

Kalman Linear Attention: Parallel Bayesian Filtering For Efficient Language Modelling and State Tracking

ResearchDGX agent

arXiv:2602.10743v2 Announce Type: replace Abstract: State-space language models such as Mamba and gated linear attention (GLA) offer linear-complexity, parallelisable alternatives to transformers, but

Model-Based and Data-Driven Hierarchical Control and Topology Co-Design for Robust Networked Systems

Local AiDGX agent

arXiv:2606.11596v1 Announce Type: cross Abstract: In this paper, we consider a class of networked systems comprising an interconnected set of linear subsystems, disturbance inputs, and performance out

Physics-Driven Spatiotemporal Modeling for AI-Generated Video Detection

ResearchDGX agent

arXiv:2510.08073v2 Announce Type: replace Abstract: AI-generated videos have achieved near-perfect visual realism (e.g., Sora), urgently necessitating reliable detection mechanisms. However, detecting

Towards Deep Learning Surrogate for the Forward Problem in Electrocardiology: A Scalable Alternative to Physics-Based Models

ResearchDGX agent

arXiv:2512.13765v2 Announce Type: replace-cross Abstract: The forward problem in electrocardiology, computing body surface potentials from cardiac electrical activity, is traditionally solved using ph

Where Do Backdoors Live? A Component-Level Analysis of Backdoor Propagation in Speech Language Models

ResearchDGX agent

arXiv:2510.01157v4 Announce Type: replace Abstract: Speech language models (SLMs) are systems of systems: independent components that unite to achieve a common goal. Despite their heterogeneous nature

10 Jun 2026

A Navigable Manifold of Hypothesized Consciousness-Spectrum States in Language Model Representations

Local AiDGX agent

arXiv:2606.09894v1 Announce Type: cross Abstract: Across contemplative, philosophical, and psychological accounts, human consciousness is often described along a similar spectrum, ranging from reactiv

A Unified Adaptive Feature Composition Framework for Multi-Task Generalization in Wireless Foundation Models

ResearchDGX agent

arXiv:2606.10277v1 Announce Type: new Abstract: Though wireless foundation models (WFMs) have shown strong potential in learning universal channel representations, their adaptation to various downstre

Anthropic has chosen the *opposite* of the safe path: they are allowing themselves, the current top lab, to use their top model for frontier…

TutorialsDGX agent

Anthropic has chosen the *opposite* of the safe path: they are allowing themselves, the current top lab, to use their top model for frontier AI research. They've said they'll sabotage others who try.

CAST: Counterfactual Labels Improve Instruction Following in Vision-Language-Action Models

ResearchDGX agent

arXiv:2508.13446v2 Announce Type: replace Abstract: Generalist robots should be able to understand and follow user instructions. Despite providing a powerful architecture for mapping open-vocabulary l

Coding agents break when models are 'almost' bug-free. But almost valid JSON is just not the same valid JSON. Fun piece here from @akshay_pa…

ToolsDGX agent

Coding agents break when models are 'almost' bug-free. But almost valid JSON is just not the same valid JSON. Fun piece here from @akshay_pachaar shows why SFT can't fix this, and how GRPO trains agai

Democratising Camera Trap AI: An Open-Source Model for Detecting UK Mammals

ResearchDGX agent

arXiv:2606.10940v1 Announce Type: cross Abstract: Camera traps have become a cornerstone of biodiversity monitoring, but the artificial intelligence that turns vast quantities of images into usable ec

Dynamics of Adversarial Attacks on Large Language Model-Based Search Engines

ResearchDGX agent

arXiv:2501.00745v3 Announce Type: replace-cross Abstract: The increasing integration of Large Language Model (LLM) based search engines has transformed the landscape of information retrieval. However,

Easy solution to slow down recursive AI self improvement: - The lab with the top-ranked model must agree THEY must not use it for working on…

TutorialsDGX agent

Easy solution to slow down recursive AI self improvement: - The lab with the top-ranked model must agree THEY must not use it for working on frontier AI - But everyone else should have access to it. B

Exact Functional ANOVA Decomposition for Categorical Inputs Models

ResearchDGX agent

arXiv:2603.02673v2 Announce Type: replace-cross Abstract: Functional ANOVA offers a principled framework for interpretability by decomposing a model's prediction into main effects and higher-order int

← Previous
1…209210211212213…1017
Next →