AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,565 results
7 Jul 2026

Gemma 4 Technical Report

Model ReleasesDGX agent

arXiv:2607.02770v1 Announce Type: cross Abstract: We introduce Gemma 4, a new generation of open-weight, natively multimodal language models in the Gemma model family. Designed to advance compute effi

GeoFlow: Geo-Aware Modeling of Inter-Area Relationships in Origin-Destination Flow Prediction and Generation

ResearchDGX agent

arXiv:2607.05257v1 Announce Type: new Abstract: Origin-destination (OD) flow modeling underpins urban planning and mobility analysis, but prevailing graph-based methods often neglect salient geographi

Hierarchical Sparse Attention Done Right: Toward Infinite Context Modeling

ResearchDGX agent

arXiv:2607.02980v1 Announce Type: cross Abstract: Scaling modern large language models (LLMs) to long contexts is limited by the quadratic computation cost, and poor length extrapolation of dense atte

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Hugging Face Models on Foundry Managed Compute

ToolsDGX agent

This article describes how to use Hugging Face models with Microsoft's Foundry Managed Compute infrastructure, likely covering integration steps, deployment options, and the benefits of running open-s

Labeled-Data-Free Meta-Learning: Efficient Task Generation Using Pre-trained Models and Unlabeled Data

ApplicationsDGX agent

arXiv:2607.02850v1 Announce Type: new Abstract: Meta-learning without labeled data is crucial for real-world applications, where obtaining labeled datasets can be expensive or restricted due to privac

Language Models Represent and Transform Concepts with Shared Geometry

ResearchDGX agent

arXiv:2607.04525v1 Announce Type: cross Abstract: How concepts are represented in neural networks is a fundamental question in machine learning. The dominant view treats concept representations as sta

Latent Clarity: Bridging World-Model Kinematics to Semantic Manifolds for Video Anomaly Anticipation

Model ReleasesDGX agent

arXiv:2607.03558v1 Announce Type: cross Abstract: Continuous video anomaly detection is dominated by reactive Multiple Instance Learning (MIL) that collapses spatiotemporal features into scalar scores

LBR: Towards Mitigating Length Bias in Large Language Models for Recommendation

SafetyDGX agent

arXiv:2607.04270v1 Announce Type: cross Abstract: Large language models (LLMs) have recently emerged as powerful backbones for recommender systems by reformulating recommendation as a token-level gene

Look Before You Leap: Distilling Tree Search into Action Evaluation for Frozen VLA Models

ResearchDGX agent

arXiv:2607.03751v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models acquire broad embodied capabilities through large-scale pretraining, yet their generalization remains far more fragi

Moonstone: A Multimodal Foundation Model and Benchmark for Lunar Remote Sensing

Model ReleasesDGX agent

arXiv:2607.03644v1 Announce Type: cross Abstract: Decades of orbital missions have produced multi-modal remote sensing data for the Moon, spanning optical imagery, spectroscopy, thermal emission, rada

Optimizing Large Language Models for Causality Assessment in Pharmacovigilance: Developing a Performance Metric as Objective for Bayesian Hyperparameter Optimization

Model ReleasesDGX agent

arXiv:2607.03704v1 Announce Type: new Abstract: Background: Growing individual case safety report (ICSR) volumes have intensified demand for scalable automated causality assessment. Large Language Mod

PRISM3D: Probabilistic Refinement and Robust Initialization for Physically Consistent Scene Modeling under Extreme Motion Blur

Model ReleasesDGX agent

arXiv:2607.03855v1 Announce Type: new Abstract: We address the inverse problem of blind 3D scene reconstruction from extremely motion-blurred images, a scenario where traditional Structure-from-Motion

RLIE: Rule Generation with Logistic Regression, Iterative Refinement, and Evaluation for Large Language Models

TutorialsDGX agent

arXiv:2510.19698v3 Announce Type: replace Abstract: Large Language Models (LLMs) can propose rules in natural language, sidestepping the need for a predefined predicate space in traditional rule learn

Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain

Model ReleasesDGX agent

arXiv:2510.17801v2 Announce Type: replace-cross Abstract: Building robots that can perceive, reason, and act in dynamic, unstructured environments remains a central challenge. Recent embodied systems

RoboVista: Evaluating Vision Language Models for Diverse Robot Applications

Model ReleasesDGX agent

arXiv:2607.04610v1 Announce Type: new Abstract: Diverse applications for robotics, such as industry and agriculture, require robots to operate across various embodiments, changing visual conditions, a

Tensor-Train Joint Modeling for Few-Step Discrete Diffusion

SafetyDGX agent

arXiv:2607.03788v1 Announce Type: new Abstract: Discrete diffusion promises orders-of-magnitude faster generation than autoregressive (AR) models for sequential discrete data, yet its full potential o

Token Communications: A Large Model-Driven Framework for Cross-modal Context-aware Semantic Communications

TutorialsDGX agent

arXiv:2502.12096v5 Announce Type: replace-cross Abstract: In this paper, we introduce token communications (TokCom), a large model-driven framework to leverage cross-modal context information in gener

ToolFailBench: Diagnosing Tool-Use Failures in LLM Agents

Model ReleasesDGX agent

arXiv:2607.04686v1 Announce Type: cross Abstract: Tool calling is central to modern language model agents, but aggregate benchmark scores often hide where tool use fails. A model that never calls a ne

TORINO: Token Reduction via Interpretable Concept Overlap in Vision-Language Models

ResearchDGX agent

arXiv:2607.04593v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) have demonstrated impressive capabilities across different tasks, but their computational cost is dominated by the large

Toward Trustworthy Large Language Model Agents in Healthcare

Model ReleasesDGX agent

arXiv:2607.05055v1 Announce Type: new Abstract: Healthcare appointment scheduling remains a persistent operational bottleneck, driven by manual coordination, fragmented legacy systems, and high admini

Towards transferable lightweight neuromorphic computing through a model-free temporal-switch framework

Model ReleasesDGX agent

arXiv:2607.02608v1 Announce Type: cross Abstract: Lightweight neuromorphic computing offers a promising route to efficient AI, with particular benefits for resource-constrained edge deployments. Howev

Video Generation Models Are Inherent Lighting Estimators

ResearchDGX agent

arXiv:2607.04674v1 Announce Type: new Abstract: Recovering dynamic environment maps from a single in-the-wild video is crucial for photorealistic rendering, yet remains a challenge. Recent video gener

5 Jul 2026

Alex Karp, frontier models and the real fight for Enterprise AI

ApplicationsDGX agent

Palantir Technologies Inc. Chief Executive Alex Karp’s recent broadside against the frontier model vendors put a knife to the throat of the central enterprise artificial intelligence debate. Karp’s ar

3 Jul 2026

ACID: Action Consistency via Inverse Dynamics for Planning with World Models

ResearchDGX agent

arXiv:2607.02403v1 Announce Type: cross Abstract: Decision-time planning with action-conditioned world models has become a popular paradigm for embodied control. However, the standard planning cost ju

BALF: Budgeted Activation-Aware Low-Rank Factorization for Fine-Tuning-Free Model Compression

Model ReleasesDGX agent

arXiv:2509.25136v3 Announce Type: replace Abstract: Activation-aware low-rank factorization techniques yield strong compression results but are generally confined to linear layers, while existing whit

Can Language Models Actually Retrieve In-Context? Drowning in Documents at Million Token Scale

ResearchDGX agent

arXiv:2607.01538v1 Announce Type: new Abstract: Language models (LMs) raise an intriguing alternative to vector-based retrieval: conditioning on an in-context corpus and directly generating a relevant

ESC: Emotional Self-Correction for Reliable Vision-Language Models

SafetyDGX agent

arXiv:2607.02089v1 Announce Type: cross Abstract: Vision-language models (VLMs) have achieved strong performance across diverse multimodal tasks, yet they remain vulnerable to unreliable reasoning. Ex

First builds of Model Y Long Wheelbase at Giga Texas

IndustryDGX agent

Tesla has begun initial production of the Model Y Long Wheelbase variant at its Gigafactory Texas facility. The Long Wheelbase version offers extended interior space and cargo capacity compared to the

From Monolingual to Multilingual: Evaluating Mamba for ASR in South African Languages

Model ReleasesDGX agent

arXiv:2607.01502v1 Announce Type: new Abstract: Recent advances in automatic speech recognition (ASR) have explored different sequence models, including Conformer-based models and newer state space mo

Language Models as Measurement Apparatus for Culture

AgentsDGX agent

arXiv:2607.02459v1 Announce Type: new Abstract: Language models are increasingly used to quantify cultural phenomena, but what makes such measurement distinctively cultural? This paper argues that NLP

Probing Chemical Language Models: Effects of Pre-training and Fine-tuning

ResearchDGX agent

arXiv:2607.02140v1 Announce Type: new Abstract: Chemical language models (CLMs) are trained with linearized representations such as SMILES, yet it remains unclear which chemically meaningful substruct

Visually Grounded Self-Reflection for Vision-Language Models via Reinforcement Learning

TutorialsDGX agent

arXiv:2607.02490v1 Announce Type: new Abstract: Large vision-language models can reason over multimodal inputs by generating textual chains of thought (CoT). A key capability exhibited in CoT reasonin

2 Jul 2026

ABot-M0.5: Unified Mobility-and-Manipulation World Action Model

Local AiDGX agent

arXiv:2607.00678v1 Announce Type: new Abstract: Mobile manipulation is a key capability for general-purpose robots, yet remains challenging for current embodied learning methods. VLA policies are typi

AI sovereignty isn’t optional. Don’t give away your alpha so easily. Protect it as much as you can. Open source models are critical and shou…

ResearchDGX agent

AI sovereignty isn’t optional. Don’t give away your alpha so easily. Protect it as much as you can. Open source models are critical and should be an important part of any individual’s, organisation’s,

But yikes does Fable write text that sounds like a parody of a Claude model on overdrive.

Model ReleasesDGX agent

This post critiques Fable AI's text generation style, suggesting it produces overly verbose or exaggerated outputs that parody Claude's characteristic writing patterns taken to an extreme. The comment

CAT: Confidence-Adaptive Thinking for Efficient Reasoning of Large Reasoning Models

ResearchDGX agent

arXiv:2607.00862v1 Announce Type: cross Abstract: Large Reasoning Models (LRMs) have achieved remarkable success on complex tasks by leveraging long chain-of-thought (CoT) trajectories, yet they frequ

Characterizing and Identifying Separable Graphical Models

ResearchDGX agent

arXiv:2607.01057v1 Announce Type: cross Abstract: We study a broad class of graphical models whose independencies correspond to vertex separation in mixed graphs with directed, undirected, and bidirec

Closed-loop coupling of personalised and foundation models for real-time treatment guidance with MRI

ResearchDGX agent

arXiv:2607.00500v1 Announce Type: cross Abstract: Image-guided therapies, including radiotherapy, biopsy and deep brain stimulation, rely on real-time targeting of anatomical structures. However, in t

Cross-Domain Generalization Failure in Lightweight Intrusion Detection Models for IIoT Networks

ResearchDGX agent

arXiv:2607.00553v1 Announce Type: cross Abstract: Lightweight machine learning models are increasingly proposed for intrusion detection in Industrial Internet of Things (IIoT) networks due to their su

Entropy-Regularized Probabilistic Gates for Sparse Model Discovery in Scarce-Data Federated Learning

Model ReleasesDGX agent

arXiv:2607.00275v1 Announce Type: cross Abstract: Federated Learning (FL) is a distributed machine learning (ML) paradigm with collaboration among multiple clients without sharing data. FL is challeng

Explainability in mulimodal deep transformation models for stroke outcome prediction

ResearchDGX agent

arXiv:2504.06299v2 Announce Type: replace-cross Abstract: Multimodal prediction models based on imaging and clinical data are increasingly used for clinical decision support, yet their interpretabilit

Generative Model Proposal based Particle Filtering for Data Assimilation

TutorialsDGX agent

arXiv:2607.01012v1 Announce Type: new Abstract: Data assimilation models state dynamics conditioned on sequential observations, and has wide-ranging scientific applications. In the filtering setting,

Korzhinskii-Net: Physics-Informed Neural Network for Sub-Surface Mineral Prospectivity Modelling

Local AiDGX agent

arXiv:2606.13695v2 Announce Type: replace-cross Abstract: Mineral prospectivity modelling (MPM) underpins exploration economics, yet most operational pipelines reduce to data-driven classifiers traine

Learn Once, Edit Anywhere: Visual Direction Transfer for Diffusion Models

TutorialsDGX agent

arXiv:2403.19645v2 Announce Type: replace Abstract: The rapid advancement of diffusion models has enabled the generation of high-fidelity images from textual prompts, yet achieving precise, disentangl

LLVM-Bench: Benchmarking and Advancing Large Language Models for LLVM Compiler Issue Resolution

Model ReleasesDGX agent

arXiv:2607.00700v1 Announce Type: cross Abstract: LLVM is a widely used compiler infrastructure whose scale and complexity make issue resolution labor-intensive and challenging. Although large languag

Lots of people are advocating for more American open-source models these days which is amazing but very few people do anything about it! Lat…

IndustryDGX agent

Lots of people are advocating for more American open-source models these days which is amazing but very few people do anything about it! Latest example, Alex Karp came out advocating for American open

MineRobot: An Actuator-Centered Kinematic Modeling and Solving Framework for Underground Mining Robots

ResearchDGX agent

arXiv:2603.22055v2 Announce Type: replace-cross Abstract: Underground mining robots are increasingly modeled for planning, operator training, and digital-twin workflows, where reliable actuator-level

Seed2.0 Model Card: Towards Intelligence Frontier for Real-World Complexity

ApplicationsDGX agent

arXiv:2607.00248v1 Announce Type: new Abstract: We present Seed2.0, a model series that takes a meaningful step toward solving complex, real-world tasks. Our approach begins with identifying users' ge

The Illusion of High Utility in Safety Alignment of Text-to-Image Diffusion Models

SafetyDGX agent

arXiv:2607.00402v1 Announce Type: cross Abstract: Safety alignment of text-to-image (T2I) diffusion models aims to suppress harmful generations while preserving utility on benign prompts. Recent metho

Training-Free Debiasing of Diffusion Models via CLIP-Guided Denoising Optimization

SafetyDGX agent

arXiv:2607.00817v1 Announce Type: new Abstract: Text-to-image diffusion models achieve impressive visual quality, yet demographic bias remains a challenge, as neutral prompts consistently produce ster

Understanding Why Language Models Hallucinate: Testing Reasoning Against Priors

SafetyDGX agent

arXiv:2607.00447v1 Announce Type: new Abstract: Large language models often produce hallucinated answers that violate prompt-level constraints. A key diagnostic question is whether these failures refl

why are the open tools so low in the list? We need to improve integration between open platforms and open models @steipete @thdxr @Teknium @…

AgentsDGX agent

why are the open tools so low in the list? We need to improve integration between open platforms and open models @steipete @thdxr @Teknium @badlogicgames! Coding agents are real users of the @huggingf

1 Jul 2026

A Large-Language-Model Supported Personalized Driving Framework for Lane Change in Highway Scenarios

Model ReleasesDGX agent

arXiv:2606.31483v1 Announce Type: new Abstract: Personalized driving can improve the user acceptance of automated driving systems. However, existing methods still provide limited support for translati

A Realistic Protocol for Evaluation of Weakly Supervised Object Localization

Model ReleasesDGX agent

arXiv:2404.10034v3 Announce Type: replace Abstract: Weakly Supervised Object Localization (WSOL) allows training deep learning models for classification and localization (LOC) using only global class-

Efficient Sampling with Discrete Diffusion Models: Sharp and Adaptive Guarantees

ResearchDGX agent

arXiv:2602.15008v2 Announce Type: replace Abstract: Diffusion models over discrete spaces have recently shown striking empirical success, yet their theoretical foundations remain incomplete. In this p

Falsification, Not Exposure: An Internally Preregistered Placebo-Controlled Decomposition of Self-Repair Feedback in Frozen Small Code Models

ResearchDGX agent

arXiv:2606.31511v1 Announce Type: cross Abstract: In deployment settings where retraining is infeasible, small frozen code models are routinely asked to repair a failed program after seeing their own

FeRA: Frequency-Energy Constrained Routing for Effective Diffusion Adaptation Fine-Tuning

Model ReleasesDGX agent

arXiv:2511.17979v2 Announce Type: replace Abstract: Diffusion models have achieved remarkable success in generative modeling, yet how to effectively adapt large pretrained models to new tasks remains

Finite Difference Flow Optimization for RL Post-Training of Text-to-Image Models

SafetyDGX agent

arXiv:2603.12893v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) has become a standard technique for post-training diffusion-based image synthesis models, as it enables learning f

Look But Don't Touch with Sparse Autoencoders for Unlearning in Diffusion Models

Local AiDGX agent

arXiv:2606.31699v1 Announce Type: cross Abstract: Sparse autoencoders (SAEs) have recently been proposed as interpretable tools for concept-level manipulation, under the assumption that isolated featu

Mixture-of-Control: State-Aware Fine-Tuning for Transformer-based Models

Model ReleasesDGX agent

arXiv:2606.31397v1 Announce Type: cross Abstract: State-based fine-tuning has emerged as a compelling alternative to weight-based adaptation for transformers, updating lightweight controls into states

← Previous
1…151152153154155…1010
Next →