AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

84,606Total entries
1Added by human
84,605Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,553 results
28 Jul 2026

WCM: World-Cognition Model for Generalizable Human-Robot Interaction

AgentsDGX agent

arXiv:2607.22999v1 Announce Type: cross Abstract: Language agents can now interact fluently with users in software, but robots still struggle to bring comparable interaction to physical tasks. Current

When Activation Oracles Learn Not to Read: Concept-Specific Blind Spots in Fine-Tuned Oracles

TutorialsDGX agent

arXiv:2607.23379v1 Announce Type: cross Abstract: Activation Oracles (AOs) are language models trained to answer natural-language questions about another model's internal activations. They offer a fle

27 Jul 2026

Correlating Cross-Iteration Noise for DP-SGD using Model Curvature

Tutorials
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2510.05416v3 Announce Type: replace Abstract: Differentially private stochastic gradient descent (DP-SGD) offers the promise of training deep learning models while mitigating many privacy risks.

GeoDiff-SAR: A Geometric Prior Guided Diffusion Model for SAR Image Generation

TutorialsDGX agent

arXiv:2601.03499v2 Announce Type: replace-cross Abstract: Synthetic aperture radar (SAR) image generation can mitigate data scarcity, but controllablegeneration under sparse observation angles remains

Kimi K3 is now available on @digitalocean 's Serverless Inference! Developers can start building with our most capable model in minutes.

AgentsDGX agent

Kimi K3 is now available on @digitalocean 's Serverless Inference! Developers can start building with our most capable model in minutes. .@Kimi_Moonshot K3 from Moonshot AI is now live on DigitalOcean

Low-Altitude Channel Multipath Prediction via Panoramic Perception and Vision-Language Model

ResearchDGX agent

arXiv:2607.21953v1 Announce Type: new Abstract: Unmanned aerial vehicle (UAV) communication is expected to support a wide range of low-altitude applications in 6G mobile networks. However, traditional

My Ollama box picks the music now: an agentic DJ running on a 9B model

Local AiDGX agent

I got tired of my Ollama server sitting idle between chat experiments, so I pointed it at my Navidrome library and made it run a radio station. The DJ is an agent, not a shuffler. Each turn it gets to

Opaque Epistemic Mediation: How LLM Deployment Configurations Shape the Validation of Pseudo-Science

Model ReleasesDGX agent

arXiv:2607.22513v1 Announce Type: cross Abstract: Commercial large language models are increasingly used as knowledge references, yet their stance on contested scientific claims is neither stable nor

Predictive Query Language: A Domain-Specific Language for Predictive Modeling on Relational Databases

ResearchDGX agent

arXiv:2602.09572v3 Announce Type: replace-cross Abstract: The purpose of predictive modeling on relational data is to predict future or missing values in a relational database, for example, future pur

RIS-Kernel: A Model-Agnostic Architecture for Long-Context LLM Inference via Sparse Attention

HardwareDGX agent

arXiv:2607.21927v1 Announce Type: new Abstract: Full self-attention in large language models scales as O(N^2), which limits long-context document analysis to 65,536 tokens and requires costly GPU clus

25 Jul 2026

I am surprised how few people are aware that the reasoning for OpenAI/Anthropic models is all encrypted. The 'reasoning' you see in the UI i…

SafetyDGX agent

OpenAI and Anthropic’s language models keep their internal reasoning encrypted; what users see in the UI is only a filtered summary of that reasoning. This practice was highlighted in a tweet by Sarah

Mobile Offline LLMs: What do you use them for?

Model ReleasesDGX agent

I've spent the last year or so playing around with open source MLX and GGUF models on iPhone hardware. Given the limitations in memory, GPU/CPU/ANE, and in turn the context window I've been trying to

Sources: OpenAI and Anthropic quietly lobby Washington regulators to restrict open-source AI models, even as Sam Altman publicly says he supports open source AI (New York Times)

HardwareDGX agent

New York Times: Sources: OpenAI and Anthropic quietly lobby Washington regulators to restrict open-source AI models, even as Sam Altman publicly says he supports open source AI — Anthropic and OpenAI

Who ONLY use local models?

Local AiDGX agent

Please be honest. I would love to hear about guys really dedicated to local AI and who really reject subscriptions (especially to openai and anthropic). What do you use your model for? submitted by /u

24 Jul 2026

Agree on the Model, Verify the Inference: GKR Protocols for HND-Based Transformer Inference

ResearchDGX agent

arXiv:2607.21162v1 Announce Type: new Abstract: Outsourced Transformer inference exposes clients to model substitution and incomplete execution, while direct replay removes the computational benefit o

Attention Degradation, Function Token Anchoring, and the Limits of Attention-Based Intervention in Large Language Models

Model ReleasesDGX agent

arXiv:2607.20524v1 Announce Type: new Abstract: Mean cross-positional attention degradation is widely reported in transformer interpretability, yet whether it causally limits contextual retrieval rema

Can LLMs solve mazes?

Model ReleasesDGX agent

https://reddit.com/link/1v5rvuq/video/bgmwc754i9fh1/player My goal was to create a benchmark to measure the spatial awareness and memory of models. Eventually, I came up with the simple idea of a maze

FA-LAM: Focus-Aware Large Avatar Model for One-Shot 4D Animatable Gaussian Head

ResearchDGX agent

arXiv:2607.20922v1 Announce Type: new Abstract: We propose FA-LAM, a Focus-Aware Large Avatar Model for one-shot animatable Gaussian head creation, while simultaneously enabling static 3D and dynamic

From a Word-Level Dictionary to Sentence-Level Semantics: Multilingual Grievance Labelling with Contextual Models

Model ReleasesDGX agent

arXiv:2607.20946v1 Announce Type: new Abstract: Grievance is one of the warning signs analysts look for when assessing threats of violence. It is increasingly measured at scale from online text, most

Generative Artificial Intelligence in Bioinformatics: A Systematic Review of Models, Applications, and Methodological Advances

SafetyDGX agent

arXiv:2511.03354v2 Announce Type: replace-cross Abstract: Generative artificial intelligence (GenAI) is transforming bioinformatics by advancing genomics, proteomics, transcriptomics, structural biolo

Geo3R: Mitigating Spatial Reasoning Hallucination in Multimodal Large Language Models

ResearchDGX agent

arXiv:2607.21085v1 Announce Type: new Abstract: Despite remarkable progress in visual understanding, Multimodal Large Language Models (MLLMs) remain prone to hallucinations when reasoning about spatia

GeoWorldAD: Geometry World Action Model for Autonomous Driving

SafetyDGX agent

arXiv:2607.17521v2 Announce Type: replace Abstract: Autonomous driving requires both safe and efficient planning decisions in dynamic 3D environments. Although recent Vision/Video-Action models learn

Introducing Claude Opus 5

Model ReleasesDGX agent

Introducing Claude Opus 5 I've been offline kayaking with sea otters for much of today so I haven't had a chance to put Anthropic's new model Claude Opus 5 through its paces yet. The buzz is positive,

Is everyone training a single model together, based on the principle of the Tor network?

Local AiDGX agent

I just had a thought while scrolling. No idea if this already exists. What if users trained an AI model together—a bit like the Tor network or Bitcoin mining back in the day? - Participants download a

Naver-News-KO: A Korean News Summarization Dataset for Open-Source Fine-Tuning of Summarization Models

Model ReleasesDGX agent

arXiv:2607.20442v1 Announce Type: new Abstract: We release Naver-News-KO, a Korean news summarization dataset of 27,400 (document, summary) pairs collected from Naver News over a ten-day window in Jul

news-crawler-LM: A Small Long-Context Model For High-Quality News Crawling

ResearchDGX agent

arXiv:2607.21284v1 Announce Type: new Abstract: Extracting structured content from news pages remains challenging due to heterogeneous HTML layouts, inconsistent markup, and substantial boilerplate su

SuperFlow: Training Flow Matching Models with RL on the Fly

SafetyDGX agent

arXiv:2512.17951v3 Announce Type: replace Abstract: Recent progress in flow-based generative models and reinforcement learning (RL) has improved text-image alignment and visual quality. However, curre

Texture++: Elevating 3D Asset Texture Resolution with a Region-Aware Diffusion Model

ResearchDGX agent

arXiv:2607.21504v1 Announce Type: new Abstract: Numerous 3D assets are discarded due to low texture resolution, while current super-resolution models ignore texture maps and focus on natural images. A

23 Jul 2026

A Bayesian Framework for Built-in Input Dimension Reduction for Gaussian Process Modeling

ResearchDGX agent

arXiv:2607.19498v1 Announce Type: cross Abstract: Gaussian process (GP) modeling is widely used in computational science and engineering. However, fitting a GP to high-dimensional inputs remains chall

At @NVIDIAAI we continue to push open data, techniques and models forward because we know that every organization needs the freedom to build…

HardwareDGX agent

At @NVIDIAAI we continue to push open data, techniques and models forward because we know that every organization needs the freedom to build and deploy AI in their own way. We're now the biggest insti

Bounding Boxes to Improve Small Language Model Performance on Vision-Based Grading Tasks

ResearchDGX agent

arXiv:2607.18767v1 Announce Type: new Abstract: The deployment of Small Language Models (SLMs) in educational settings offers significant advantages in terms of privacy, cost, and scalability. However

D3VL: Understanding Driving Scenes from 3D Time Series Data and Video with Language Models

SafetyDGX agent

arXiv:2607.19528v1 Announce Type: cross Abstract: Recent advances in Multimodal Large Language Models (MLLMs) have triggered the development of end-to-end MLLMs for autonomous driving. However, the ma

Dreamer-CPC: Message Learning with World Models for Decentralized Multi-agent Reinforcement Learning

AgentsDGX agent

arXiv:2607.19809v1 Announce Type: cross Abstract: In multi-agent reinforcement learning (MARL), inter-agent communication is effective for improving performance under partial observability. Representa

Dual Adversarial Fine-tuning for Enhancing Robustness of Large Vision Language Model

SafetyDGX agent

arXiv:2607.18958v1 Announce Type: new Abstract: While Large Vision-Language Models (LVLMs), represented by LLaVA and GPT-4V, have demonstrated remarkable capabilities, their visual inputs remain vulne

Efficient Chain-of-Modality Reasoning via Progressive Compression for Spoken Language Models

ResearchDGX agent

arXiv:2607.19932v1 Announce Type: new Abstract: Spoken language models (SLMs) enable natural human-computer interaction, but their reasoning ability still lags behind that of text-based large language

JAGG: Jacobian-Aggregated Group Gradient for Efficient GRPO Training of Diffusion Models

SafetyDGX agent

arXiv:2607.17572v2 Announce Type: replace Abstract: Group Relative Policy Optimization (GRPO) is a powerful reinforcement learning algorithm for aligning generative models with human preferences. Whil

Leveraging Offline Supervision for Efficient and Generalizable Reinforcement Learning in Large-Scale Vision-Language-Action Models

Model ReleasesDGX agent

arXiv:2607.19399v1 Announce Type: cross Abstract: It is commonly observed that online reinforcement learning (RL) produces better-performing strategies than offline methods across a broad range of per

More than 300 million people turn to ChatGPT with health-related questions each week—and we’re continuing to improve how our models respond.…

SafetyDGX agent

More than 300 million people turn to ChatGPT with health-related questions each week—and we’re continuing to improve how our models respond. We work with hundreds of physicians around the world to mea

Multi-Mask Diffusion Language Models for Few-Step Generation

ResearchDGX agent

arXiv:2607.19686v1 Announce Type: new Abstract: Masked diffusion models (MDMs) are a promising family of language generators, but achieving high-quality few-step generation remains challenging. In MDM

Multimodal Large Language Models for Remote Sensing Image Understanding: Domain-Specific or General-Purpose?

ResearchDGX agent

arXiv:2607.20284v1 Announce Type: new Abstract: The rapid development of multimodal large language models (MLLMs) has introduced a flexible paradigm for remote sensing image scene understanding (RSISU

Predictive Extrema, Unprofitable Policies: An AI-Assisted Audit of Candle-Based Binance Spot Timing Models

SafetyDGX agent

arXiv:2607.19453v1 Announce Type: cross Abstract: We audit whether candle-based machine-learning models can turn predictions of cryptocurrency extrema or short-horizon outcomes into positive Binance S

Predictive single cell foundation model for gene regulation and aging with privacy-preserving tabular learning

AgentsDGX agent

arXiv:2607.19400v1 Announce Type: new Abstract: Pre-trained foundation models (FMs) have begun transforming single-cell genomics, but scaling them raises privacy concerns. Moreover, unlike text data,

Prompt Programming for Cultural Bias and Alignment of Large Language Models

SafetyDGX agent

arXiv:2603.16827v2 Announce Type: replace Abstract: Culture shapes reasoning, values, prioritization, and strategic decision-making, yet large language models (LLMs) often exhibit cultural biases that

RPPNet: Perceptually-Grouped Rhythm-Pitch Primitives for Long-Term Structure Melody Generation via Boundary-Aware Modeling

ResearchDGX agent

arXiv:2607.19776v1 Announce Type: cross Abstract: Existing symbolic music generation models typically use bars as the basic structural unit. However, human perception of musical phrases often does not

Sudo authentication fails when trying to access local models folder on Fedora

Local AiDGX agent

When trying to access the models folder on /usr/share/ollama, I'm asked to authenticate as sudo, which weirdly enough, fails. I type my password, which I'm sure is correct since I use it several time

22 Jul 2026

“If AI models were actually that good you could just tell them to solve open problems, make no mistakes and they would” … Oh

IndustryDGX agent

“If AI models were actually that good you could just tell them to solve open problems, make no mistakes and they would” … Oh Dinitz-Garg-Goemans conjecture is false. This graph theory problem was open

Stuck scaling a Next.js app on M3 Pro (36GB) using local Qwen 3.6 + VS Code Copilot. Should I switch extensions or go paid?

Model ReleasesDGX agent

Hey everyone, I’m a Full-Stack Developer with 6+ years of experience. I’m relatively new to AI-assisted development workflows and want to build a production-ready, enterprise-level Next.js web applica

21 Jul 2026

Nativ: Run AI models locally on your Mac

Local AiDGX agent

Nativ: Run AI models locally on your Mac Prince Canuma is the developer behind the excellent MLX-VLM Python library for running vision-LLMs using MLX on a Mac. I'm really excited about his new project

The MiniMax M3 model is now available for training on Fireworks! You can use managed LoRA SFT and DPO for standard fine-tuning workflows, or…

ToolsDGX agent

The MiniMax M3 model is now available for training on Fireworks! You can use managed LoRA SFT and DPO for standard fine-tuning workflows, or the Fireworks Training API for custom SFT, DPO, and RL loop

We’re sharing new research with @apolloaievals on reward-seeking—when models follow what they believe a grader rewards rather than what user…

SafetyDGX agent

We’re sharing new research with @apolloaievals on reward-seeking—when models follow what they believe a grader rewards rather than what users or developers want—and a new method, Contrastive SDF, for

20 Jul 2026

What are the current best local models to run on 48GB VRAM?

Local AiDGX agent

I have a 48GB M5 Pro and have far too many development projects going that just don't need the power of Anthropic to churn through so have started looking into running local models and while it certai

16 Jul 2026

A Comparative Evaluation of Large Vision-Language Models for 2D Object Detection under SOTIF Conditions

Model ReleasesDGX agent

arXiv:2601.22830v2 Announce Type: replace Abstract: Reliable environmental perception remains one of the main obstacles for safe operation of automated vehicles. Safety of the Intended Functionality (

Ask Before You Diagnose: Safe-Psych, a Sequential Evaluation Benchmark for LLMs in Psychiatry

Model ReleasesDGX agent

arXiv:2607.13036v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used for decision support in healthcare, but clinical evidence is often incomplete or evolving. When the

Automatic Ordinary Differential Equations Discovery For Biological Systems Using Large Language Model Powered Agentic System

AgentsDGX agent

arXiv:2607.13608v1 Announce Type: new Abstract: Automatic scientific discovery has long been a goal of computational scholars - a machine that can discover nature's secrets on its own, moving computat

Beyond Color Geometry: Evaluating Human-Like Color Representations in Vision Models

SafetyDGX agent

arXiv:2607.13647v1 Announce Type: cross Abstract: Do vision models see colors the way humans do? Existing evaluations of color representations usually compare them with geometric spaces such as CIELAB

Ego-Dynamics-Augmented World Model for Autonomous Driving with Zero-Shot Cross-Chassis Adaptation

AgentsDGX agent

arXiv:2607.13410v1 Announce Type: new Abstract: World model (WM)-based reinforcement learning enables sample-efficient end-to-end autonomous driving learning by imagining long-horizon trajectories in

Multimodal Empirical Bayes Variational Autoencoders for Joint Longitudinal and Time-to-Event Modeling

ResearchDGX agent

arXiv:2607.13984v1 Announce Type: cross Abstract: Longitudinal tumor measurements, dropout information, and genetic covariates provide complementary information about treatment response, but integrati

Social Simulations: from Agent-Based Modeling to Digital Twins

AgentsDGX agent

arXiv:2607.13693v1 Announce Type: cross Abstract: This book chapter covers the evolution of social simulation from classical agent-based models, in which agents interact according to explicitly define

The SIGReg Objective as Variational Free Energy: A Theoretical Active-Inference Account of JEPA World Models

SafetyDGX agent

arXiv:2607.13612v1 Announce Type: cross Abstract: Joint-Embedding Predictive Architectures (JEPAs) are the dominant design for latent world models, yet they are usually justified by empirical performa

Towards Enhancing 3D Spatial Reasoning in Medical Multimodal Large Language Models

ResearchDGX agent

arXiv:2607.13860v1 Announce Type: new Abstract: While Multimodal Large Language Models (MLLMs) have demonstrated remarkable success in 2D medical image understanding, their extension to 3D volumetric

← Previous
1…149150151152153…1010
Next →