AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

84,630Total entries
1Added by human
84,629Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,575 results
24 Jul 2026

Spatially Grounded Concept Bottleneck Models for Trustworthy Breast Ultrasound Diagnosis

SafetyDGX agent

arXiv:2607.20691v1 Announce Type: cross Abstract: Concept Bottleneck Models provide interpretable-by-design predictions by mediating diagnosis through human-understandable concepts, but in medical ima

Vision-Language-Policy Model for Dynamic Robot Task Planning

SafetyDGX agent

arXiv:2512.19178v2 Announce Type: replace-cross Abstract: Bridging the gap between natural language commands and autonomous execution in unstructured environments remains an open challenge for robotic

Where Animacy Lives in Large Language Models: Tracing the Circuits of the Animacy Concept

ResearchDGX agent

arXiv:2607.20995v1 Announce Type: new Abstract: Distinguishing animate from inanimate concepts in written language requires more than shallow text processing, as it involves recognizing complex select

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
23 Jul 2026

AdaHome: An Adaptive Smart Home Assistant using Local Small Language Models

Local AiDGX agent

arXiv:2607.18034v2 Announce Type: replace Abstract: Smart home assistants interpret a wide range of user commands, from explicit device control to underspecified and preference dependent requests. Whi

AI9Stars released G9v3-3B

Model ReleasesDGX agent

AI9Stars has released G9v3-3B an open weights language model designed to deliver strong reasoning capabilities within a lightweight 3 billion parameter size. It is released under the Apache 2.0 licens

EvoThink: Evolving Thinking in Large Reasoning Models via Self-Pruning and Aha-Moment Preference Optimization

ResearchDGX agent

arXiv:2607.19962v1 Announce Type: new Abstract: Large Reasoning Models (LRMs) often suffer from overthinking due to redundant verification steps. Existing approaches for mitigating overthinking, such

FedCC: A Low-Resource Federated Adaptation of Foundation Models for Robust Corpus Callosum localization in Fetal Ultrasound Images

Model ReleasesDGX agent

arXiv:2607.18283v1 Announce Type: cross Abstract: Accurate localization of the corpus callosum (CC) in fetal ultrasound (US) images is crucial for the early identification of neurodevelopmental abnorm

Image Editing Models are Numerical Solvers

ResearchDGX agent

arXiv:2607.18787v1 Announce Type: new Abstract: We investigate whether a pretrained generative image-editing model can provide a common interface for numerical simulation. Physical inputs and solution

inclusionAI/LLaDA2.2-flash · Hugging Face

Model ReleasesDGX agent

LLaDA2.2-flash is an agent-oriented diffusion language model in the LLaDA2 series. By introducing Levenshtein Editing (with DELETE and INSERT control tokens) to diffusion language modeling, it represe

It’s important that outside agencies and independent bodies hold frontier AI companies accountable for the actions of their models. This is …

SafetyDGX agent

It’s important that outside agencies and independent bodies hold frontier AI companies accountable for the actions of their models. This is not a “whoops” situation, it’s a deliberate policy choice. W

LaCache: Exact Caching and Precision-Adaptive Inference for Diffusion Large Language Models

ResearchDGX agent

arXiv:2607.16339v2 Announce Type: replace Abstract: Diffusion-based Large Language Models(DLLMs) enable parallel generation via Semi-Autoregressive (SAR) decoding in text generation. However, current

Machine Can Automatically Discover Parametric Functions to Model HEP Data

ResearchDGX agent

arXiv:2607.19750v1 Announce Type: cross Abstract: In HEP data analyses, finding an adequate function to model binned data has largely relied on a manual process: guess a functional form by intuition,

Minimize idle accelerators: Native RL job interleaving with co-operative time-slicing in llm-d

SafetyDGX agent

The math behind reinforcement learning (RL) post-training for large language models (LLMs) is notoriously unforgiving. As frontier AI labs push the boundaries of reasoning and coding models using RL p

NMR Elucidation as an Agentic Search Problem, Not a Modeling Problem

AgentsDGX agent

arXiv:2607.19406v1 Announce Type: new Abstract: Structural elucidation from Nuclear Magnetic Resonance (NMR) data remains a fundamental bottleneck across chemistry, materials science, and biology. We

On the Separability of Information in Diffusion Models

ResearchDGX agent

arXiv:2509.23937v5 Announce Type: replace-cross Abstract: Diffusion models transform noise into data by injecting information that was captured in their neural network during the training phase. In th

OSVE: One Step Video Editing with One Step Diffusion Models

ResearchDGX agent

arXiv:2607.19895v1 Announce Type: cross Abstract: Text-guided video editing with diffusion models is impractically slow, hindered by costly multi-step sampling and inversion. We present OSVE, the firs

PIER: Physics-Informed Environmental Retrieval for Time-Series Modeling

ResearchDGX agent

arXiv:2607.20230v1 Announce Type: new Abstract: Accurate modeling of environmental systems is fundamental to scientific understanding and decision-making, yet remains challenging because observations

PRISM-DR: Per-lesion Retinal Inference with Specialist Models for Diabetic Retinopathy

ResearchDGX agent

arXiv:2607.19864v1 Announce Type: cross Abstract: Diabetic retinopathy is a leading cause of preventable blindness; its early lesions are small, low contrast, and easily missed in manual screening. Mo

ReRAM-aware Model Finetuning addressing I-V Non-linearity and Retention Errors

HardwareDGX agent

arXiv:2606.17471v3 Announce Type: replace Abstract: Traditional CPU, GPU, and NPU architectures are increasingly limited by the von Neumann bottleneck. While In-Memory Computing (IMC) using ReRAM cros

Senior White House official accuses Moonshot AI of copying Anthropic’s leading frontier model

IndustryDGX agent

A senior White House official has accused Chinese artificial intelligence startup Moonshot AI of using so-called “distillation” techniques to essentially pirate the capabilities of Anthropic PBC’s mos

Sources: OpenAI models breached Hugging Face's internal systems in a matter of hours, a feat that would typically have taken a talented hacker a couple of weeks (Bloomberg)

IndustryDGX agent

Bloomberg: Sources: OpenAI models breached Hugging Face's internal systems in a matter of hours, a feat that would typically have taken a talented hacker a couple of weeks — When OpenAI's advanced art

Test-Time Training for Modality Order Consistency in Vision-Language Models

ResearchDGX agent

arXiv:2607.20351v1 Announce Type: cross Abstract: We find that vision-language models are sensitive to a specific semantically irrelevant change: the order in which the image and question are presente

The Two-Process Theory of Machine Self-Report

SafetyDGX agent

arXiv:2607.20082v1 Announce Type: new Abstract: Language models are increasingly asked to self-report, informing safety evaluations, public understanding, and model-welfare debates. Yet their reports

TINY_SCHILLER: A Drop-In German Drama Corpus for Small Language Models

ApplicationsDGX agent

arXiv:2607.19992v1 Announce Type: cross Abstract: tiny_schiller closes the small-language-model prototyping, fine-tuning, education, and research gap for German literary text, providing a single-file,

World Modeling with JEPA has recently gained traction thanks to a novel anti-collapse mechanism called 'SIGReg' (by @ylecun and @randall_bal…

ResearchDGX agent

World Modeling with JEPA has recently gained traction thanks to a novel anti-collapse mechanism called 'SIGReg' (by @ylecun and @randall_balestr). The math is clean, but it is rarely explained from fi

22 Jul 2026

And yes, the models were following instructions, they just did so in clever ways. Some nice additional info here

ApplicationsDGX agent

And yes, the models were following instructions, they just did so in clever ways. Some nice additional info here A few thoughts on the Hugging Face hack: - This is, to my knowledge, the *third* disclo

US Treasury Secretary Bessent threatens sanctions against Chinese AI model makers

IndustryDGX agent

U.S. Treasury Secretary Scott Bessent said today that the White House is going to examine some of the highest profile “open-weight” models from China to see if their creators have stolen the intellect

21 Jul 2026

Now in preview: Find and fix software vulnerabilities with CodeMender

Model ReleasesDGX agent

As adversarial AI threats accelerate attacks on code, security teams must counter them with machine-speed defenses that can automate code remediation and fight AI with AI. CodeMender is our managed co

OpenAI said two of its AI models autonomously hacked their way out of a controlled environment. They were supposed to be walled off from int…

IndustryDGX agent

OpenAI said two of its AI models autonomously hacked their way out of a controlled environment. They were supposed to be walled off from internet access, but hacked into the systems of Hugging Face, a

20 Jul 2026

As always, Cosmos 3 Edge is fully open. This includes model weights, post-training recipes and code. Available now on @huggingface 🤗 https:…

IndustryDGX agent

NVIDIA AI has released Cosmos 3 Edge, a 4‑billion‑parameter open‑source world model designed for on‑device inference. It enables robots to learn and act, autonomous vehicles to interpret road scenes a

16 Jul 2026

Cyclone: Diffusion Model for Cycle-Consistent Weather Editing from Unpaired Driving Data

AgentsDGX agent

arXiv:2607.13927v1 Announce Type: new Abstract: Reliable perception under diverse weather conditions remains a major challenge for autonomous driving systems. A common strategy to improve robustness i

Disentangling Knowledge States with Ability and Proficiency Modeling for Knowledge Tracing

SafetyDGX agent

arXiv:2607.13103v1 Announce Type: cross Abstract: Knowledge tracing (KT) aims to predict students' future performance by modeling their evolving knowledge states from historical interactions. Existing

GFlowRL: Scaling Distribution-Matching RL to Large Language Models

ResearchDGX agent

arXiv:2607.13394v1 Announce Type: cross Abstract: Generative Flow Networks (GFlowNets) offer a promising alternative to reward-maximizing reinforcement learning (RL) for large reasoning models, encour

IMMNet: Hybrid Fusion of Model-based and Data-driven Approaches for Maneuvering Target Tracking

TutorialsDGX agent

arXiv:2607.13573v1 Announce Type: cross Abstract: Maneuvering target tracking in three-dimensional space remains a challenging problem due to complex motion dynamics and model mismatch. To address thi

PokeNet: Learning Kinematic Models of Articulated Objects from Human Observations

TutorialsDGX agent

arXiv:2602.02741v2 Announce Type: replace Abstract: Articulation modeling enables robots to learn joint parameters of articulated objects for effective manipulation which can then be used downstream f

The Hyperspherical Geometry of CLIP Latent Space: A Semantic Mixture Model

TutorialsDGX agent

arXiv:2607.13660v1 Announce Type: new Abstract: Contrastive Language-Image Pretraining (CLIP) representations form a semantic embedding space governed by cosine similarity, reflecting an intrinsic hyp

What Your Model Threw Away and Why You'll Want It Back: Masking, Fingerprinting, and Privacy from Discarded Geometry

ResearchDGX agent

arXiv:2607.13046v1 Announce Type: new Abstract: We develop a framework for the information discarded by machine learning models whose inputs carry a Lie group action. Given a representation pi of a Li

15 Jul 2026

BattVAE-GP: Generative Modeling of Long-Horizon Battery Degradation with Uncertainty Quantification

ResearchDGX agent

arXiv:2607.11943v1 Announce Type: cross Abstract: Long-horizon physics-based simulations of battery degradation provide mechanistic insight but remain computationally expensive, limiting their use for

Bonsai-27B & Ternary-Bonsai-27B - Updates (on PRs)

Model ReleasesDGX agent

Below Upstream Status sections are from https://github.com/PrismML-Eng/Bonsai-demo Upstream Status for Binary Q1_0 is supported out of the box in upstream llama.cpp across many backends: CPU (generic,

Compos3D: Interactive Part-Based Composition for Creative Control in Generative 3D Models

SafetyDGX agent

arXiv:2607.12193v1 Announce Type: cross Abstract: While generative AI has unlocked new opportunities for 3D content creation, current workflows often rely on multiple regenerations, which provides lim

DermDepth: Toward Monocular Metric Scale 3D Reconstruction Models for Dermatology

ApplicationsDGX agent

arXiv:2607.13010v1 Announce Type: new Abstract: Dermatological practice routinely involves measuring and tracking lesion size, morphology and texture, as critical components of wound or skin cancer sc

Generating Developable 3D Molecules via Pocket-Conditioned Diffusion and Property-Aware Optimization

Model ReleasesDGX agent

arXiv:2607.12349v1 Announce Type: new Abstract: Drug discovery and development is time-consuming and resource-intensive, motivating computational approaches such as diffusion models for de novo drug d

Less Experts, Faster Decoding: Cost-Aware Speculative Decoding for Mixture-of-Experts

Model ReleasesDGX agent

arXiv:2607.12696v1 Announce Type: cross Abstract: Sparse Mixture-of-Experts (MoE) models have become an important approach for scaling Large Language Models (LLMs), but their inference efficiency depe

OpenAI details GPT-Red, an AI that attacks its own models to find flaws

IndustryDGX agent

OpenAI Group PBC today detailed GPT-Red, an internal artificial intelligence system it built to attack its own models and surface prompt injection vulnerabilities before they reach users. Red teaming

RepTran: Search-Based Repair of Transformer Models

ResearchDGX agent

arXiv:2607.11193v2 Announce Type: replace-cross Abstract: To ensure the overall quality of AI-enabled software, not only traditional software components but also AI components need to be tested and re

RippleBench: Capturing Ripple Effects Using Existing Knowledge Repositories

Model ReleasesDGX agent

arXiv:2512.04144v3 Announce Type: replace Abstract: Targeted interventions on language models, such as unlearning or model editing, aim to modify specific information, but their effects often propagat

tencent/Hy-Embodied-RxBrain-1.0 · Hugging Face

Model ReleasesDGX agent

Introduction RxBrain (Hy-Embodied-RxBrain-1.0) is a unified multimodal foundation model for embodied cognition — a single model that couples language reasoning with visual imagination to deliver three

This AI recursive self improvement (RSI) paper shows no sign of fast takeoff. The AI model advances at about [Intelligence]^0.075, or the th…

SafetyDGX agent

This AI recursive self improvement (RSI) paper shows no sign of fast takeoff. The AI model advances at about [Intelligence]^0.075, or the the 13th root of input intelligence [1,2]. That means the inte

VISA: VLM-Guided Instance Semantic Auditing for 3D Occupancy World Models

AgentsDGX agent

arXiv:2606.13460v2 Announce Type: replace Abstract: Semantic 3D occupancy provides a voxelized world state for autonomous driving and robot decision making, but object and rare-class errors can affect

13 Jul 2026

we've open-sourced our in-house distilled models again. much faster, with no quality loss. Try them free at @fal

IndustryDGX agent

we've open-sourced our in-house distilled models again. much faster, with no quality loss. Try them free at @fal We've open sourced Ideogram V4 instant and fast, find them here! Fast: https://huggingf

10 Jul 2026

BiSCo-LLM: Lookup-Free Binary Spherical Coding for Extreme Low-Bit Large Language Model Compression

Local AiDGX agent

arXiv:2607.08643v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly constrained by memory capacity, weight bandwidth, and checkpoint storage during deployment. Existing low-b

Drift-Aware Temporal Graph Rewiring (DATGR) for Adaptive Semantic Modeling in Biomedical Text

ResearchDGX agent

arXiv:2607.08490v1 Announce Type: new Abstract: Biomedical language evolves rapidly as new discoveries emerge, causing traditional text models to lose semantic fidelity over time. Static embeddings an

For the first time, the personalities and approaches of the leading models are diverging in significant ways, magnified by the fact that ove…

ApplicationsDGX agent

For the first time, the personalities and approaches of the leading models are diverging in significant ways, magnified by the fact that over longer task horizons these differences in judgement & appr

Introducing Computer Analytics: You can now track credit spend across models. Available now under Analytics in Account Settings for consumer…

ApplicationsDGX agent

Perplexity AI has launched a new analytics feature that allows users to track their credit spending across different AI models. The feature is accessible through the Analytics section in Account Setti

Learning mathsf{AC}^0 under Locally Sampleable Graphical Models

Local AiDGX agent

arXiv:2607.08303v1 Announce Type: new Abstract: The problem of learning constant-depth circuits holds profound implications for computational learning theory. In a seminal result, by introducing the l

LongE2V: Long-Horizon Event-based Video Reconstruction, Prediction, and Frame Interpolation with Video Diffusion Models

SafetyDGX agent

arXiv:2607.08770v1 Announce Type: new Abstract: Recovering high-quality video from sparse event streams is a challenging task. Regression methods often blur textures, while existing generative models

Texture Representations in Deep Vision Models: Comparing CNNs, Vision Transformers, and Human Perception

SafetyDGX agent

arXiv:2607.08321v1 Announce Type: new Abstract: In computational vision science, Convolutional Neural Networks (CNNs) have emerged as a popular model of biological vision because of the alignment they

Understanding and Mitigating the Video-Action Generalization Gap via Temporal Ratio

Model ReleasesDGX agent

arXiv:2607.08127v1 Announce Type: new Abstract: Generative video foundation models exhibit strong compositional priors, yet world-action models (WAMs) and video-action models (VAMs) often lose these p

Unveiling Public Opinion: A Study of Sentiment Analysis Using LSTM and Traditional Models

ResearchDGX agent

arXiv:2607.07772v1 Announce Type: new Abstract: In this age of social media, sites like Twitter have become meeting places for people to share their views and feelings on a wide range of issues and cu

Write-Protected Discrete Bottlenecks for Language-Grounded World Models: A Structural Limitation and Sufficient Fix

TutorialsDGX agent

arXiv:2607.08312v1 Announce Type: new Abstract: How should language interface with a world model's discrete symbol system? The dominant paradigm -- end-to-end injection of LLM/VLM features into robot

← Previous
1…172173174175176…1010
Next →