AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
84,661Total entries
1Added by human
84,660Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,596 results
Model Releases

Attention Degradation, Function Token Anchoring, and the Limits of Attention-Based Intervention in Large Language Models

DGX agent

arXiv:2607.20524v1 Announce Type: new Abstract: Mean cross-positional attention degradation is widely reported in transformer interpretability, yet whether it causally limits contextual retrieval rema

model-releasesarxiv-cs-ai
24 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Can LLMs solve mazes?

DGX agent

https://reddit.com/link/1v5rvuq/video/bgmwc754i9fh1/player My goal was to create a benchmark to measure the spatial awareness and memory of models. Eventually, I came up with the simple idea of a maze

model-releasesr-localllama
24 Jul 2026
Research

FA-LAM: Focus-Aware Large Avatar Model for One-Shot 4D Animatable Gaussian Head

DGX agent

arXiv:2607.20922v1 Announce Type: new Abstract: We propose FA-LAM, a Focus-Aware Large Avatar Model for one-shot animatable Gaussian head creation, while simultaneously enabling static 3D and dynamic

researcharxiv-cs-cv
24 Jul 2026
Model Releases

From a Word-Level Dictionary to Sentence-Level Semantics: Multilingual Grievance Labelling with Contextual Models

DGX agent

arXiv:2607.20946v1 Announce Type: new Abstract: Grievance is one of the warning signs analysts look for when assessing threats of violence. It is increasingly measured at scale from online text, most

model-releasesarxiv-cs-cl
24 Jul 2026
Safety

Generative Artificial Intelligence in Bioinformatics: A Systematic Review of Models, Applications, and Methodological Advances

DGX agent

arXiv:2511.03354v2 Announce Type: replace-cross Abstract: Generative artificial intelligence (GenAI) is transforming bioinformatics by advancing genomics, proteomics, transcriptomics, structural biolo

safetyarxiv-cs-ai
24 Jul 2026
Research

Geo3R: Mitigating Spatial Reasoning Hallucination in Multimodal Large Language Models

DGX agent

arXiv:2607.21085v1 Announce Type: new Abstract: Despite remarkable progress in visual understanding, Multimodal Large Language Models (MLLMs) remain prone to hallucinations when reasoning about spatia

researcharxiv-cs-cv
24 Jul 2026
Safety

GeoWorldAD: Geometry World Action Model for Autonomous Driving

DGX agent

arXiv:2607.17521v2 Announce Type: replace Abstract: Autonomous driving requires both safe and efficient planning decisions in dynamic 3D environments. Although recent Vision/Video-Action models learn

safetyarxiv-cs-ro
24 Jul 2026
Model Releases

Introducing Claude Opus 5

DGX agent

Introducing Claude Opus 5 I've been offline kayaking with sea otters for much of today so I haven't had a chance to put Anthropic's new model Claude Opus 5 through its paces yet. The buzz is positive,

model-releasessimon-willison
24 Jul 2026
Local Ai

Is everyone training a single model together, based on the principle of the Tor network?

DGX agent

I just had a thought while scrolling. No idea if this already exists. What if users trained an AI model together—a bit like the Tor network or Bitcoin mining back in the day? - Participants download a

local-air-ollama
24 Jul 2026
Model Releases

Naver-News-KO: A Korean News Summarization Dataset for Open-Source Fine-Tuning of Summarization Models

DGX agent

arXiv:2607.20442v1 Announce Type: new Abstract: We release Naver-News-KO, a Korean news summarization dataset of 27,400 (document, summary) pairs collected from Naver News over a ten-day window in Jul

model-releasesarxiv-cs-cl
24 Jul 2026
Research

news-crawler-LM: A Small Long-Context Model For High-Quality News Crawling

DGX agent

arXiv:2607.21284v1 Announce Type: new Abstract: Extracting structured content from news pages remains challenging due to heterogeneous HTML layouts, inconsistent markup, and substantial boilerplate su

researcharxiv-cs-cl
24 Jul 2026
Safety

SuperFlow: Training Flow Matching Models with RL on the Fly

DGX agent

arXiv:2512.17951v3 Announce Type: replace Abstract: Recent progress in flow-based generative models and reinforcement learning (RL) has improved text-image alignment and visual quality. However, curre

safetyarxiv-cs-cv
24 Jul 2026
Research

Texture++: Elevating 3D Asset Texture Resolution with a Region-Aware Diffusion Model

DGX agent

arXiv:2607.21504v1 Announce Type: new Abstract: Numerous 3D assets are discarded due to low texture resolution, while current super-resolution models ignore texture maps and focus on natural images. A

researcharxiv-cs-cv
24 Jul 2026
Research

A Bayesian Framework for Built-in Input Dimension Reduction for Gaussian Process Modeling

DGX agent

arXiv:2607.19498v1 Announce Type: cross Abstract: Gaussian process (GP) modeling is widely used in computational science and engineering. However, fitting a GP to high-dimensional inputs remains chall

researcharxiv-cs-lg
23 Jul 2026
Hardware

At @NVIDIAAI we continue to push open data, techniques and models forward because we know that every organization needs the freedom to build…

DGX agent

At @NVIDIAAI we continue to push open data, techniques and models forward because we know that every organization needs the freedom to build and deploy AI in their own way. We're now the biggest insti

hardwareclem-delangue--x
23 Jul 2026
Research

Bounding Boxes to Improve Small Language Model Performance on Vision-Based Grading Tasks

DGX agent

arXiv:2607.18767v1 Announce Type: new Abstract: The deployment of Small Language Models (SLMs) in educational settings offers significant advantages in terms of privacy, cost, and scalability. However

researcharxiv-cs-cv
23 Jul 2026
Safety

D3VL: Understanding Driving Scenes from 3D Time Series Data and Video with Language Models

DGX agent

arXiv:2607.19528v1 Announce Type: cross Abstract: Recent advances in Multimodal Large Language Models (MLLMs) have triggered the development of end-to-end MLLMs for autonomous driving. However, the ma

safetyarxiv-cs-ai
23 Jul 2026
Agents

Dreamer-CPC: Message Learning with World Models for Decentralized Multi-agent Reinforcement Learning

DGX agent

arXiv:2607.19809v1 Announce Type: cross Abstract: In multi-agent reinforcement learning (MARL), inter-agent communication is effective for improving performance under partial observability. Representa

agentsarxiv-cs-lg
23 Jul 2026
Safety

Dual Adversarial Fine-tuning for Enhancing Robustness of Large Vision Language Model

DGX agent

arXiv:2607.18958v1 Announce Type: new Abstract: While Large Vision-Language Models (LVLMs), represented by LLaVA and GPT-4V, have demonstrated remarkable capabilities, their visual inputs remain vulne

safetyarxiv-cs-cv
23 Jul 2026
Research

Efficient Chain-of-Modality Reasoning via Progressive Compression for Spoken Language Models

DGX agent

arXiv:2607.19932v1 Announce Type: new Abstract: Spoken language models (SLMs) enable natural human-computer interaction, but their reasoning ability still lags behind that of text-based large language

researcharxiv-cs-cl
23 Jul 2026
Safety

JAGG: Jacobian-Aggregated Group Gradient for Efficient GRPO Training of Diffusion Models

DGX agent

arXiv:2607.17572v2 Announce Type: replace Abstract: Group Relative Policy Optimization (GRPO) is a powerful reinforcement learning algorithm for aligning generative models with human preferences. Whil

safetyarxiv-cs-lg
23 Jul 2026
Model Releases

Leveraging Offline Supervision for Efficient and Generalizable Reinforcement Learning in Large-Scale Vision-Language-Action Models

DGX agent

arXiv:2607.19399v1 Announce Type: cross Abstract: It is commonly observed that online reinforcement learning (RL) produces better-performing strategies than offline methods across a broad range of per

model-releasesarxiv-cs-ai
23 Jul 2026
Safety

More than 300 million people turn to ChatGPT with health-related questions each week—and we’re continuing to improve how our models respond.…

DGX agent

More than 300 million people turn to ChatGPT with health-related questions each week—and we’re continuing to improve how our models respond. We work with hundreds of physicians around the world to mea

safetyopenai--x
23 Jul 2026
Research

Multi-Mask Diffusion Language Models for Few-Step Generation

DGX agent

arXiv:2607.19686v1 Announce Type: new Abstract: Masked diffusion models (MDMs) are a promising family of language generators, but achieving high-quality few-step generation remains challenging. In MDM

researcharxiv-cs-cl
23 Jul 2026
Research

Multimodal Large Language Models for Remote Sensing Image Understanding: Domain-Specific or General-Purpose?

DGX agent

arXiv:2607.20284v1 Announce Type: new Abstract: The rapid development of multimodal large language models (MLLMs) has introduced a flexible paradigm for remote sensing image scene understanding (RSISU

researcharxiv-cs-cv
23 Jul 2026
Safety

Predictive Extrema, Unprofitable Policies: An AI-Assisted Audit of Candle-Based Binance Spot Timing Models

DGX agent

arXiv:2607.19453v1 Announce Type: cross Abstract: We audit whether candle-based machine-learning models can turn predictions of cryptocurrency extrema or short-horizon outcomes into positive Binance S

safetyarxiv-cs-ai
23 Jul 2026
Agents

Predictive single cell foundation model for gene regulation and aging with privacy-preserving tabular learning

DGX agent

arXiv:2607.19400v1 Announce Type: new Abstract: Pre-trained foundation models (FMs) have begun transforming single-cell genomics, but scaling them raises privacy concerns. Moreover, unlike text data,

agentsarxiv-cs-lg
23 Jul 2026
Safety

Prompt Programming for Cultural Bias and Alignment of Large Language Models

DGX agent

arXiv:2603.16827v2 Announce Type: replace Abstract: Culture shapes reasoning, values, prioritization, and strategic decision-making, yet large language models (LLMs) often exhibit cultural biases that

safetyarxiv-cs-ai
23 Jul 2026
Research

RPPNet: Perceptually-Grouped Rhythm-Pitch Primitives for Long-Term Structure Melody Generation via Boundary-Aware Modeling

DGX agent

arXiv:2607.19776v1 Announce Type: cross Abstract: Existing symbolic music generation models typically use bars as the basic structural unit. However, human perception of musical phrases often does not

researcharxiv-cs-ai
23 Jul 2026
Local Ai

Sudo authentication fails when trying to access local models folder on Fedora

DGX agent

When trying to access the models folder on /usr/share/ollama, I'm asked to authenticate as sudo, which weirdly enough, fails. I type my password, which I'm sure is correct since I use it several time

local-air-ollama
23 Jul 2026
Industry

“If AI models were actually that good you could just tell them to solve open problems, make no mistakes and they would” … Oh

DGX agent

“If AI models were actually that good you could just tell them to solve open problems, make no mistakes and they would” … Oh Dinitz-Garg-Goemans conjecture is false. This graph theory problem was open

industryemad-mostaque--x
22 Jul 2026
Model Releases

Stuck scaling a Next.js app on M3 Pro (36GB) using local Qwen 3.6 + VS Code Copilot. Should I switch extensions or go paid?

DGX agent

Hey everyone, I’m a Full-Stack Developer with 6+ years of experience. I’m relatively new to AI-assisted development workflows and want to build a production-ready, enterprise-level Next.js web applica

model-releasesr-ollama
22 Jul 2026
Local Ai

Nativ: Run AI models locally on your Mac

DGX agent

Nativ: Run AI models locally on your Mac Prince Canuma is the developer behind the excellent MLX-VLM Python library for running vision-LLMs using MLX on a Mac. I'm really excited about his new project

local-aisimon-willison
21 Jul 2026
Tools

The MiniMax M3 model is now available for training on Fireworks! You can use managed LoRA SFT and DPO for standard fine-tuning workflows, or…

DGX agent

The MiniMax M3 model is now available for training on Fireworks! You can use managed LoRA SFT and DPO for standard fine-tuning workflows, or the Fireworks Training API for custom SFT, DPO, and RL loop

toolsfireworks-ai--x
21 Jul 2026
Safety

We’re sharing new research with @apolloaievals on reward-seeking—when models follow what they believe a grader rewards rather than what user…

DGX agent

We’re sharing new research with @apolloaievals on reward-seeking—when models follow what they believe a grader rewards rather than what users or developers want—and a new method, Contrastive SDF, for

safetyopenai--x
21 Jul 2026
Local Ai

What are the current best local models to run on 48GB VRAM?

DGX agent

I have a 48GB M5 Pro and have far too many development projects going that just don't need the power of Anthropic to churn through so have started looking into running local models and while it certai

local-air-ollama
20 Jul 2026
Model Releases

A Comparative Evaluation of Large Vision-Language Models for 2D Object Detection under SOTIF Conditions

DGX agent

arXiv:2601.22830v2 Announce Type: replace Abstract: Reliable environmental perception remains one of the main obstacles for safe operation of automated vehicles. Safety of the Intended Functionality (

model-releasesarxiv-cs-cv
16 Jul 2026
Model Releases

Ask Before You Diagnose: Safe-Psych, a Sequential Evaluation Benchmark for LLMs in Psychiatry

DGX agent

arXiv:2607.13036v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used for decision support in healthcare, but clinical evidence is often incomplete or evolving. When the

model-releasesarxiv-cs-ai
16 Jul 2026
Agents

Automatic Ordinary Differential Equations Discovery For Biological Systems Using Large Language Model Powered Agentic System

DGX agent

arXiv:2607.13608v1 Announce Type: new Abstract: Automatic scientific discovery has long been a goal of computational scholars - a machine that can discover nature's secrets on its own, moving computat

agentsarxiv-cs-ai
16 Jul 2026
Safety

Beyond Color Geometry: Evaluating Human-Like Color Representations in Vision Models

DGX agent

arXiv:2607.13647v1 Announce Type: cross Abstract: Do vision models see colors the way humans do? Existing evaluations of color representations usually compare them with geometric spaces such as CIELAB

safetyarxiv-cs-ai
16 Jul 2026
Agents

Ego-Dynamics-Augmented World Model for Autonomous Driving with Zero-Shot Cross-Chassis Adaptation

DGX agent

arXiv:2607.13410v1 Announce Type: new Abstract: World model (WM)-based reinforcement learning enables sample-efficient end-to-end autonomous driving learning by imagining long-horizon trajectories in

agentsarxiv-cs-ro
16 Jul 2026
Research

Multimodal Empirical Bayes Variational Autoencoders for Joint Longitudinal and Time-to-Event Modeling

DGX agent

arXiv:2607.13984v1 Announce Type: cross Abstract: Longitudinal tumor measurements, dropout information, and genetic covariates provide complementary information about treatment response, but integrati

researcharxiv-cs-lg
16 Jul 2026
Agents

Social Simulations: from Agent-Based Modeling to Digital Twins

DGX agent

arXiv:2607.13693v1 Announce Type: cross Abstract: This book chapter covers the evolution of social simulation from classical agent-based models, in which agents interact according to explicitly define

agentsarxiv-cs-ai
16 Jul 2026
Safety

The SIGReg Objective as Variational Free Energy: A Theoretical Active-Inference Account of JEPA World Models

DGX agent

arXiv:2607.13612v1 Announce Type: cross Abstract: Joint-Embedding Predictive Architectures (JEPAs) are the dominant design for latent world models, yet they are usually justified by empirical performa

safetyarxiv-cs-ai
16 Jul 2026
Research

Towards Enhancing 3D Spatial Reasoning in Medical Multimodal Large Language Models

DGX agent

arXiv:2607.13860v1 Announce Type: new Abstract: While Multimodal Large Language Models (MLLMs) have demonstrated remarkable success in 2D medical image understanding, their extension to 3D volumetric

researcharxiv-cs-cv
16 Jul 2026
Model Releases

Agents-A1-4B (Qwen3.7-4B ???) : Scaling the Horizon, Not the Parameters

DGX agent

MODEL + GGUF : https://huggingface.co/InternScience/models?search=a1-4b Technical Report Benchmark Qwen3.5-4B Agents-A1-4B Qwen3.5 Qwen3.6 Nex-N2-mini Agents-A1 🧠 Dense Models (~4B) 🔀 MoE Models (35B-

model-releasesr-localllama
15 Jul 2026
Research

Entropy-Preserving Supervised Fine-Tuning via Adaptive Self-Distillation for Large Reasoning Models

DGX agent

arXiv:2602.02244v3 Announce Type: replace-cross Abstract: The standard post-training recipe for large reasoning models, supervised fine-tuning followed by reinforcement learning (SFT-then-RL), may lim

researcharxiv-cs-cl
15 Jul 2026
Local Ai

Graph Feedback Controls Consensus and Clique Formation in Open-Weight Language-Model Populations

DGX agent

arXiv:2607.12077v1 Announce Type: new Abstract: Multi-agent language-model systems increasingly route local interactions, yet the runtime interaction graph is often treated as an implementation detail

local-aiarxiv-cs-ai
15 Jul 2026
← Previous
1…187188189190191…1263
Next →