AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

84,648Total entries
1Added by human
84,647Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,588 results
23 Jun 2026

Fine-grained Human Motion Understanding with Language Models

ResearchDGX agent

arXiv:2606.20888v1 Announce Type: new Abstract: In this work, we propose methodname, an LLM-based model for fine-grained human motion understanding that represents motion as a sequence of skeletal pos

Flatness Preserves Instruction Following in Vision-Language-Action Models

ApplicationsDGX agent

arXiv:2606.23641v1 Announce Type: new Abstract: Vision-language-action (VLA) models have the potential for open-world generalization by leveraging pretrained vision-language representations, yet downs

Hierarchical Sparse Circuit Extraction from Billion-Parameter Language Models through Scalable Attribution Graph Decomposition

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2601.12879v2 Announce Type: replace Abstract: Extracting sparse circuits from billion-parameter transformers is constrained by O(2^n) search cost and pervasive feature reuse across co-active pat

Interpretable Probabilistic Medical Image Segmentation via Gaussian Process with Explicit Modelling of Annotation Bias and Variability

SafetyDGX agent

arXiv:2606.23177v1 Announce Type: new Abstract: Deep learning-based medical image segmentation models are trained using annotations that exhibit systematic bias and variability across raters. While pr

MAGNIFIED: RL Fine-tuning of Multimodal Large Language Models for Motion Planning

SafetyDGX agent

arXiv:2606.20641v1 Announce Type: cross Abstract: Multi-modal Large Language Models (MLLMs) have demonstrated remarkable capabilities in semantic understanding and common sense reasoning, making them

MambaADv2: Evolving Duality-enhanced State Space Model for Unsupervised Anomaly Detection

Local AiDGX agent

arXiv:2606.23126v1 Announce Type: new Abstract: While recent advancements in anomaly detection have demonstrated the efficacy of CNN- and Transformer-based approaches, these architectures face inheren

Measuring Model-Induced Discrimination via Efficient Fairness Approximation

SafetyDGX agent

arXiv:2405.09251v2 Announce Type: replace Abstract: Providing various machine learning (ML) applications in the real world, concerns about discrimination hidden in ML models are growing, particularly

Model-Free Robust Average-Reward Reinforcement Learning with Sample Complexity Analysis

SafetyDGX agent

arXiv:2505.12462v3 Announce Type: replace Abstract: Robust reinforcement learning (RL) under the average-reward criterion is essential for long-term decision-making, particularly when the environment

Multi-AUV Marine Life Tracking with Single Hydrophone Payloads via a Hidden Markov Model Equipped Particle Filter

Local AiDGX agent

arXiv:2606.22335v1 Announce Type: new Abstract: Researchers tag and track marine animals to study migration patterns, human impacts on behavior, and behavioral shifts due to climate change. Accurate d

PeLAP-A: Adaptive Latent Pruning for Lightweight Latent Diffusion Models

ResearchDGX agent

arXiv:2606.23086v1 Announce Type: new Abstract: Latent diffusion models achieve strong generative performance by operating in a compressed latent space produced by a variational autoencoder (VAE). How

Policy-as-Data: Learning Generalizable HOI Diffusion Models from Simulated Physics

SafetyDGX agent

arXiv:2606.22806v1 Announce Type: new Abstract: Synthesizing realistic Human-Object Interactions (HOI) is critical for creating embodied avatars and functional virtual environments. However, current d

Protocol-Aware Tokenization and Architecture Co-Design for Wireless Packet Foundation Models

ResearchDGX agent

arXiv:2606.20587v1 Announce Type: cross Abstract: What matters more for building foundation models for wireless packet traces: the tokenizer or the architecture or both? To answer this question, we bu

Read the technical paper on Krea 2 https://www.krea.ai/blog/krea-2-technical-report Download the model weights https://github.com/krea-ai/kr…

Local AiDGX agent

Krea 2 is a technical advancement in AI image generation with newly released model weights available for download on GitHub. The technical report details the improvements and capabilities of this vers

RelightAnyone: A Generalized Relightable 3D Gaussian Head Model

SafetyDGX agent

arXiv:2601.03357v2 Announce Type: replace Abstract: 3D Gaussian Splatting (3DGS) has become a standard approach to reconstruct and render photorealistic 3D head avatars. A major challenge is to religh

Revisiting OmniAnomaly for Anomaly Detection: performance metrics and comparison with PCA-based models

ResearchDGX agent

arXiv:2603.18985v2 Announce Type: replace-cross Abstract: Deep learning models have become the dominant approach for multivariate time series anomaly detection (MTSAD), often reporting substantial per

Scalable Training of Spatially Grounded 2D Vision-Language Models for Radiology

TutorialsDGX agent

arXiv:2606.20477v2 Announce Type: replace Abstract: We study how to train visually grounded vision-language models (VLMs) for radiology without manual spatial annotations. We introduce RefRad2D, a lar

Sources: the Trump administration is pressing Meta to submit its AI models for voluntary review; Meta is the only major US AI developer without an agreement (New York Times)

SafetyDGX agent

New York Times: Sources: the Trump administration is pressing Meta to submit its AI models for voluntary review; Meta is the only major US AI developer without an agreement — Federal officials are urg

Token Factory: Efficiently Integrating Diverse Signals into Large Recommendation Models

ApplicationsDGX agent

arXiv:2606.19635v2 Announce Type: replace-cross Abstract: Large Recommendation Models (LRMs) have demonstrated promising capabilities in industry-scale recommendation tasks. However, holistically inte

TooBad: Backdoor Diffusion Models with Ultra-Low Poison Rate and Imperceptible Trigger

ResearchDGX agent

arXiv:2606.23362v1 Announce Type: cross Abstract: Diffusion models (DMs), despite their impressive capabilities across a wide range of generative tasks, have been shown to be vulnerable to backdoor at

Training-Free Semantic Correction for Autoregressive Visual Models

SafetyDGX agent

arXiv:2606.22550v1 Announce Type: new Abstract: Autoregressive visual models (AVMs) based on next-scale prediction have emerged as a prominent paradigm for image and video synthesis. However, decompos

Vesta: A Generalist Embodied Reasoning Model

Local AiDGX agent

arXiv:2606.20905v1 Announce Type: new Abstract: Robots operating in open-world environments must seamlessly integrate localization, spatial reasoning, navigation, and long-horizon planning. While spec

When Does a Video-Language Model Stop Watching? Reward Strength Controls the Formation and Reversal of Visual Shortcuts in Multimodal RLVR

ResearchDGX agent

arXiv:2606.22043v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) is increasingly applied to large vision-language models (LVLMs), yet outcome-only optimization c

22 Jun 2026

Going to cross 3M public models & 1M public datasets on @huggingface in a few days. Open-source AI is on fire!

IndustryDGX agent

Hugging Face was approaching milestones of 3 million public models and 1 million public datasets on its platform, reflecting rapid growth in open-source AI resources. The announcement highlights the e

Is there a market where folk are predicting when Reflection AI will drop their first model This is probably as much compute as currently use…

IndustryDGX agent

Is there a market where folk are predicting when Reflection AI will drop their first model This is probably as much compute as currently used by all the Chinese open source companies together (more ad

My version of AI mania is when I get a spidey feeling that the models have changed. Opus 4.8 feels very different today.

IndustryDGX agent

Allie K. Miller expresses subjective observations about perceiving changes in Claude Opus 4.8's behavior or capabilities, describing an intuitive sense ('spidey feeling') that the model feels notably

Sources: Meta internally exposed data from its employee-tracking program meant to help train its AI models, including full prompts and private conversations (Wired)

IndustryDGX agent

Wired: Sources: Meta internally exposed data from its employee-tracking program meant to help train its AI models, including full prompts and private conversations — Employees had previously raised co

21 Jun 2026

MaineCoon is the first video model that focuses on social interactions: facial expressions, emotions, fluid conversation, audio-lip sync, et…

HardwareDGX agent

MaineCoon is the first video model that focuses on social interactions: facial expressions, emotions, fluid conversation, audio-lip sync, etc. Really impressive inference specs: 22B params, 47.5 FPS o

19 Jun 2026

API prices of key AI models: US vs China

IndustryDGX agent

This post likely compares the pricing of major AI model APIs between the United States and China, highlighting cost differences that reflect different market conditions, regulatory environments, and c

11 Jun 2026

Adapting Vision-Language Models from Iconic to Inclusive for Multi-Label Recognition Without Labels

SafetyDGX agent

arXiv:2606.11626v1 Announce Type: new Abstract: Understanding multi-label images remains a challenging task in computer vision. With the rapid progress of vision-language multimodal learning, vision-l

Automated Creativity Evaluation of Language Models Across Open-Ended Tasks

AgentsDGX agent

arXiv:2606.11762v1 Announce Type: cross Abstract: Large language models (LLMs) have achieved remarkable progress in language understanding, reasoning, and generation, sparking growing interest in thei

AVIS: Adaptive Test-Time Scaling for Vision-Language Models

SafetyDGX agent

arXiv:2606.11576v1 Announce Type: cross Abstract: Modern Vision-Language Models (VLMs) benefit from chain-of-thought prompting and test-time scaling, but these gains often come with prohibitive infere

Can AI Reason Like an Urban Planner? Benchmarking Large Language Models Against Professional Judgment

SafetyDGX agent

arXiv:2606.11678v1 Announce Type: new Abstract: Problem, Research Strategy, and Findings: The rise of large language models (LLMs) raises a key question for urban planning: which forms of professional

Detecting AI-Generated Content on Social Media with Multi-modal Language Models

ApplicationsDGX agent

arXiv:2606.11200v1 Announce Type: new Abstract: Generative AI has enabled the creation of photorealistic images and videos that are increasingly disseminated on social media, often used for spam, misi

Detecting Sensitive Personal Information in Japanese Pre-Training Corpora for Large Language Models

ResearchDGX agent

arXiv:2606.12114v1 Announce Type: new Abstract: Sensitive personal information can appear in large-scale pre-training corpora for large language models (LLMs). Detecting and filtering such information

Diffusion-based Cumulative Adversarial Purification for Vision Language Models

ApplicationsDGX agent

arXiv:2506.03933v2 Announce Type: replace-cross Abstract: Vision Language Models (VLMs) have shown remarkable capabilities in multimodal understanding, yet their susceptibility to adversarial perturba

Evaluating Bias in Phoneme-Based Automatic Speech Recognition Systems: An Analysis of IPA Transcription Models

SafetyDGX agent

arXiv:2606.11639v1 Announce Type: new Abstract: The popularization of automatic speech recognition (ASR) systems has increased exploration of the demographic biases related to race, age, gender, and a

I built a 100% local, CPU-only voice loop for Ollama — talk to your models hands-free (Silero VAD + Parakeet STT + Supertonic TTS 3)

Local AiDGX agent

A developer created a fully local, CPU-based voice interface for Ollama that enables hands-free conversation with AI models by combining three open-source components: Silero VAD (voice activity detect

Language Shapes Mental Health Evaluations in Large Language Models

ResearchDGX agent

arXiv:2603.06910v2 Announce Type: replace Abstract: Multilingual large language models (LLMs) are increasingly used in socially sensitive mental health contexts, including support chatbots, screening,

On the Optimal Reasoning Length for RL-Trained Language Models

ResearchDGX agent

arXiv:2602.09591v3 Announce Type: replace-cross Abstract: Reinforcement learning substantially improves reasoning in large language models, but it also tends to lengthen chain-of-thought outputs and i

Prediction-Powered Risk Monitoring of Deployed Models for Detecting Harmful Distribution Shifts

ResearchDGX agent

arXiv:2602.02229v2 Announce Type: replace Abstract: We study the problem of monitoring model performance in dynamic environments where labeled data are limited. To this end, we propose prediction-powe

Steering Where to Listen: Instruction-Based Activation Steering Redirects Temporal Attention in Large Audio-Language Models

ResearchDGX agent

arXiv:2606.11400v1 Announce Type: cross Abstract: Large Audio-Language Models (LALMs) excel at audio understanding but expose little about where in an audio signal they attend. We introduce instructio

10 Jun 2026

Access OpenAI models and Codex through your Oracle cloud commitment

ApplicationsDGX agent

OpenAI models and Codex are available to Oracle Cloud customers through an integrated partnership, allowing users to access these AI capabilities within their existing Oracle cloud infrastructure and

An LLM-Native Psychometric Instrument Does Not Predict LLM Behavior: Evidence Across 25 Models

SafetyDGX agent

arXiv:2606.09843v1 Announce Type: cross Abstract: Large language models (LLMs) produce stable self-reports on personality inventories, but these self-reports do not predict observed behavior. Whether

Attention-Discounted Adaptive Sampler for Masked Diffusion Language Models

ResearchDGX agent

arXiv:2606.10829v1 Announce Type: cross Abstract: Masked diffusion language models can reduce inference steps by revealing multiple tokens per denoising iteration, but this parallelism is fragile: pos

Blind denoising diffusion models and the blessings of dimensionality

ResearchDGX agent

arXiv:2602.09639v2 Announce Type: replace Abstract: Denoising diffusion models (DDMs) are state-of-the-art methods for learning densities from data across numerous domains, yet many aspects of the tra

Business World Model

AgentsDGX agent

arXiv:2606.10044v1 Announce Type: new Abstract: Businesses are increasingly adopting AI-enabled tools to improve productivity, reduce costs, and enhance products and services. However, the transformat

Can Multi-Agent LLMs Identify Their Peers? Stylometric Fingerprinting in Role-Constrained Political Analysis

Model ReleasesDGX agent

arXiv:2606.09854v1 Announce Type: cross Abstract: Multi-agent large language model (LLM) pipelines for political statement analysis are vulnerable to peer-preservation bias: models tend to protect pee

Deep Generative Model for Human Mobility Behavior

ResearchDGX agent

arXiv:2510.06473v3 Announce Type: replace-cross Abstract: Understanding and modeling human mobility is central to challenges in transport planning, sustainable urban design, and public health. Despite

Improving PET/CT-Based Whole-Body Lesion Segmentation Using Prediction Uncertainty-Augmented Models

Local AiDGX agent

arXiv:2606.10115v1 Announce Type: new Abstract: Accurate lesion segmentation from whole-body Positron Emission Tomography (PET)/Computed Tomography (CT) scans is essential for cancer staging and treat

Learning the Universe: Posterior Reliability of Neural Generative Models in High-Dimensional Field-Level Inference of Cosmic Initial Conditions

ApplicationsDGX agent

arXiv:2606.10023v1 Announce Type: cross Abstract: Accurate posterior estimation is central to scientific inference, as uncertainties determine what can be reliably learned from observational data. Whi

Lium raises $5.5M to unlock complex scientific data for AI models

AgentsDGX agent

Lium, a startup formerly known as Astromind, today announced the launch of an “agentic harness” that helps large language models dig into the most complex and messiest datasets. The launch comes after

Mean Flow Distillation: Robust and Stable Distillation for Flow Matching Models

SafetyDGX agent

arXiv:2606.11155v1 Announce Type: new Abstract: Flow Matching models have demonstrated strong performance across a wide range of generative tasks. However, their reliance on ODE-based iterative sampli

Mechanistic Analysis of Alignment Algorithms in Language Models

SafetyDGX agent

arXiv:2606.09850v1 Announce Type: cross Abstract: Post-training alignment algorithms are predominantly evaluated as black boxes, obscuring how they reshape language models' internal computations. We p

Null-Space Constrained Low-Rank Adaptation for Response-Specified Large Language Model Unlearning

Local AiDGX agent

arXiv:2606.10989v1 Announce Type: new Abstract: Large language model unlearning aims to suppress designated undesirable knowledge while preserving benign capabilities. Many unlearning objectives focus

Routing-Aware Expert Calibration for Machine Unlearning in Mixture-of-Experts Language Models

ResearchDGX agent

arXiv:2606.10338v1 Announce Type: cross Abstract: Machine unlearning is increasingly important for large language models, yet unlearning in Mixture-of-Experts (MoE) architectures remains underexplored

Scaling Self-Supervised Speech Models Uncovers Deep Linguistic Relationships: Evidence from the Pacific Cluster

ResearchDGX agent

arXiv:2603.07238v2 Announce Type: replace Abstract: Similarities between language representations derived from Self-Supervised Speech Models (S3Ms) have been observed to primarily reflect geographic p

SpeechJBB: Probing Safety Alignment and Comprehension in Large Audio Language Models under Code-Switched Speech

SafetyDGX agent

arXiv:2606.06037v2 Announce Type: cross Abstract: Large audio language models (LALMs) are increasingly deployed in real-world applications, yet their safety alignment is still primarily evaluated on m

TacForeSight: Force-Guided Tactile World Model for Contact-Rich Manipulation

Local AiDGX agent

arXiv:2606.11184v1 Announce Type: new Abstract: Contact-rich manipulation requires robots to continuously perceive and regulate evolving physical interactions under dynamic contact transitions or comp

Uncovering Vulnerability of Vision-Language-Action Models under Joint-Level Physical Faults

SafetyDGX agent

arXiv:2606.10501v1 Announce Type: new Abstract: Deploying Vision-Language-Action (VLA) models in real robotic systems requires robustness not only to semantic and perceptual variations, but also to em

Using Probabilistic Programs to Train Inductive Reasoning in Large Language Models

SafetyDGX agent

arXiv:2606.09856v1 Announce Type: cross Abstract: Post-training Large Language Models (LLMs) for reasoning typically focuses on deductive tasks such as mathematics and coding where correctness is veri

← Previous
1…177178179180181…1010
Next →