AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
83,745 results
13 Apr 2026

Toward Hardware-Agnostic Quadrupedal World Models via Morphology Conditioning

Model ReleasesDGX agent

arXiv:2604.08780v1 Announce Type: cross Abstract: World models promise a paradigm shift in robotics, where an agent learns the underlying physics of its environment once to enable efficient planning a

Toward World Models for Epidemiology

SafetyDGX agent

arXiv:2604.09519v1 Announce Type: new Abstract: World models have emerged as a unifying paradigm for learning latent dynamics, simulating counterfactual futures, and supporting planning under uncertai

Towards Context-Aware Image Anonymization with Multi-Agent Reasoning

AgentsDGX agent

arXiv:2603.27817v3 Announce Type: replace-cross Abstract: Street-level imagery contains personally identifiable information (PII), some of which is context-dependent. Existing anonymization methods ei

Content type
AllBlogX PostPaperYouTubeRedditGitHub

Towards developing future-ready skills with generative AI

ApplicationsDGX agent

Google Research introduces **Vantage**, a research experiment that uses generative AI to assess students' future-ready skills through simulated conversational environments, aiming to provide educators

Towards Knowledgeable Deep Research: Framework and Benchmark

Model ReleasesDGX agent

arXiv:2604.07720v2 Announce Type: replace Abstract: Deep Research (DR) requires LLM agents to autonomously perform multi-step information seeking, processing, and reasoning to generate comprehensive r

Towards Lifelong Aerial Autonomy: Geometric Memory Management for Continual Visual Place Recognition in Dynamic Environments

Model ReleasesDGX agent

arXiv:2604.09038v1 Announce Type: cross Abstract: Robust geo-localization in changing environmental conditions is critical for long-term aerial autonomy. While visual place recognition (VPR) models pe

Towards Linguistically-informed Representations for English as a Second or Foreign Language: Review, Construction and Application

ResearchDGX agent

arXiv:2604.09008v1 Announce Type: cross Abstract: The widespread use of English as a Second or Foreign Language (ESFL) has sparked a paradigm shift: ESFL is not seen merely as a deviation from standar

Towards Responsible Multimodal Medical Reasoning via Context-Aligned Vision-Language Models

SafetyDGX agent

arXiv:2604.08815v1 Announce Type: new Abstract: Medical vision-language models (VLMs) show strong performance on radiology tasks but often produce fluent yet weakly grounded conclusions due to over-re

Tracing the Chain: Deep Learning for Stepping-Stone Intrusion Detection

ResearchDGX agent

arXiv:2604.08800v1 Announce Type: cross Abstract: Stepping-stone intrusions (SSIs) are a prevalent network evasion technique in which attackers route sessions through chains of compromised intermediat

Trained a Qwen2.5-0.5B-Instruct bf16 model on Reddit post summarization task with GRPO [P]

ResearchDGX agent

A community practitioner post on r/MachineLearning documenting an experiment fine-tuning Alibaba's Qwen2.5-0.5B-Instruct model in bf16 precision on a Reddit post summarization task using GRPO (Group R

Training event-based neural networks with exact gradients via Differentiable ODE Solving in JAX

SafetyDGX agent

arXiv:2603.08146v3 Announce Type: replace Abstract: Existing frameworks for gradient-based training of spiking neural networks face a trade-off: discrete-time methods using surrogate gradients support

Training-free, Perceptually Consistent Low-Resolution Previews with High-Resolution Image for Efficient Workflows of Diffusion Models

TutorialsDGX agent

arXiv:2604.09227v1 Announce Type: cross Abstract: Image generative models have become indispensable tools to yield exquisite high-resolution (HR) images for everyone, ranging from general users to pro

Traj2Action: A Co-Denoising Framework for Trajectory-Guided Human-to-Robot Skill Transfer

SafetyDGX agent

arXiv:2510.00491v3 Announce Type: replace-cross Abstract: Learning diverse manipulation skills for real-world robots is severely bottlenecked by the reliance on costly and hard-to-scale teleoperated d

Transferable FB-GNN-MBE Framework for Potential Energy Surfaces: Data-Adaptive Transfer Learning in Deep Learned Many-Body Expansion Theory

ResearchDGX agent

arXiv:2604.09320v1 Announce Type: cross Abstract: Mechanistic understanding and rational design of complex chemical systems depend on fast and accurate predictions of electronic structures beyond indi

Trending well! We're glad the traces are useful to the @NousResearch Hermes Agent community. A third batch is in progress.

Model ReleasesDGX agent

Trending well! We're glad the traces are useful to the @NousResearch Hermes Agent community. A third batch is in progress. Very cool open-source traces from @TheZachMueller @LambdaAPI: https://hugging

Tried Ollama Cloud, just realize only Kimi model accept images

Local AiDGX agent

A Reddit user exploring Ollama Cloud noted that, at the time of their post, only the Kimi model supported image (vision/multimodal) inputs among the available cloud models. Kimi K2.5 is a native multi

TRU: Targeted Reverse Update for Efficient Multimodal Recommendation Unlearning

Model ReleasesDGX agent

arXiv:2604.02183v2 Announce Type: replace Abstract: Multimodal recommendation systems (MRS) jointly model user-item interaction graphs and rich item content, but this tight coupling makes user data di

Truncated Rectified Flow Policy for Reinforcement Learning with One-Step Sampling

SafetyDGX agent

arXiv:2604.09159v1 Announce Type: new Abstract: Maximum entropy reinforcement learning (MaxEnt RL) has become a standard framework for sequential decision making, yet its standard Gaussian policy para

Try it immediately on Ollama's cloud: ollama run gemma4:31b-cloud Try it with OpenClaw: ollama launch openclaw --model gemma4:31b-cloud Try …

Model ReleasesDGX agent

Try it immediately on Ollama's cloud: ollama run gemma4:31b-cloud Try it with OpenClaw: ollama launch openclaw --model gemma4:31b-cloud Try it with Claude Code: ollama launch claude --model gemma4:31b

TurboOCR: 270–1200 img/s OCR with Paddle + TensorRT (C++/CUDA, FP16) [P]

HardwareDGX agent

TurboOCR is a high-performance OCR project that combines PaddleOCR with NVIDIA TensorRT, implemented in C++ and CUDA, achieving throughput of 270–1,200 images per second using FP16 half-precision infe

Turning Anime into Real and testing Klein9b vs Qwen Edit 2511 (Workflow Included)

Model ReleasesDGX agent

This r/StableDiffusion post showcases a workflow for converting anime-style images into photorealistic outputs, directly comparing two AI image editing models: Klein 9B and Qwen Image Edit 2511. Qwen

TurPy: a physics-based and differentiable optical turbulence simulator for algorithmic development and system optimization

Model ReleasesDGX agent

arXiv:2604.07248v2 Announce Type: replace-cross Abstract: Developing optical systems for free-space applications requires simulation tools that accurately capture turbulence-induced wavefront distorti

U-Cast: A Surprisingly Simple and Efficient Frontier Probabilistic AI Weather Forecaster

HardwareDGX agent

arXiv:2604.09041v1 Announce Type: cross Abstract: AI-based weather forecasting now rivals traditional physics-based ensembles, but state-of-the-art (SOTA) models rely on specialized architectures and

UAV-Track VLA: Embodied Aerial Tracking via Vision-Language-Action Models

Model ReleasesDGX agent

arXiv:2604.02241v2 Announce Type: replace Abstract: Embodied visual tracking is crucial for Unmanned Aerial Vehicles (UAVs) executing complex real-world tasks. In dynamic urban scenarios with complex

UHD Low-Light Image Enhancement via Real-Time Enhancement Methods with Clifford Information Fusion

ResearchDGX agent

arXiv:2604.09321v1 Announce Type: cross Abstract: Considering efficiency, ultra-high-definition (UHD) low-light image restoration is extremely challenging. Existing methods based on Transformer archit

UIPress: Bringing Optical Token Compression to UI-to-Code Generation

ResearchDGX agent

arXiv:2604.09442v1 Announce Type: new Abstract: UI-to-Code generation requires vision-language models (VLMs) to produce thousands of tokens of structured HTML/CSS from a single screenshot, making visu

Unbiased Rectification for Sequential Recommender Systems Under Fake Orders

SafetyDGX agent

arXiv:2604.08550v1 Announce Type: cross Abstract: Fake orders pose increasing threats to sequential recommender systems by misleading recommendation results through artificially manipulated interactio

Uncertainty-Aware Transformers: Conformal Prediction for Language Models

ResearchDGX agent

arXiv:2604.08885v1 Announce Type: new Abstract: Transformers have had a profound impact on the field of artificial intelligence, especially on large language models and their variants. However, as was

Uncertainty Estimation for the Open-Set Text Classification systems

Model ReleasesDGX agent

arXiv:2604.08560v1 Announce Type: cross Abstract: Accurate uncertainty estimation is essential for building robust and trustworthy recognition systems. In this paper, we consider the open-set text cla

Unified Multimodal Uncertain Inference

Model ReleasesDGX agent

arXiv:2604.08701v1 Announce Type: new Abstract: We introduce Unified Multimodal Uncertain Inference (UMUI), a multimodal inference task spanning text, audio, and video, where models must produce calib

UniSemAlign: Text-Prototype Alignment with a Foundation Encoder for Semi-Supervised Histopathology Segmentation

SafetyDGX agent

arXiv:2604.09169v1 Announce Type: new Abstract: Semi-supervised semantic segmentation in computational pathology remains challenging due to scarce pixel-level annotations and unreliable pseudo-label s

Universal Approximation with XL MIMO Systems: OTA Classification via Trainable Analog Combining

ResearchDGX agent

arXiv:2504.12758v3 Announce Type: replace-cross Abstract: In this paper, we show that an eXtremely Large (XL) Multiple-Input Multiple-Output (MIMO) wireless system with appropriate analog combining co

Unmasking Puppeteers: Leveraging Biometric Leakage to Disarm Impersonation in AI-based Videoconferencing

ResearchDGX agent

arXiv:2510.03548v3 Announce Type: replace-cross Abstract: AI-based talking-head videoconferencing systems reduce bandwidth by sending a compact pose-expression latent and re-synthesizing RGB at the re

Update: Distilled v1.1 is live

Local AiDGX agent

This Reddit post from r/StableDiffusion announces the release of 'Distilled v1.1,' an updated version of a knowledge-distilled Stable Diffusion model, likely building on prior distillation work that p

Use ollama like the year is still 1998

Local AiDGX agent

This r/ollama Reddit post appears to be a community discussion about using Ollama in a minimal, retro, or stripped-down fashion — likely exploring terminal-based, command-line-only, or low-tech intera

Used LTX 2.3 anchor frame injection to maintain brand consistency across AI video — before/after

Local AiDGX agent

This r/StableDiffusion post demonstrates a practical workflow in which a user employs LTX 2.3's anchor frame injection technique — using keyframes such as first, middle, or last frames to condition vi

Using Synthetic Data for Machine Learning-based Childhood Vaccination Prediction in Narok, Kenya

ResearchDGX agent

arXiv:2604.08902v1 Announce Type: new Abstract: Background: Limited data utilization in low-resource settings poses a barrier to the vaccine delivery ecosystem, undermining efforts to achieve equitabl

V-CAGE: Vision-Closed-Loop Agentic Generation Engine for Robotic Manipulation

AgentsDGX agent

arXiv:2604.09036v1 Announce Type: new Abstract: Scaling Vision-Language-Action (VLA) models requires massive datasets that are both semantically coherent and physically feasible. However, existing sce

v0.20.7

Local AiDGX agent

Ollama v0.20.7 is a patch-level release in the v0.20.x series of Ollama, the open-source platform for running large language models locally. Based on the progression of the 0.20.x releases — which hav

v0.20.7-rc0: gemma4: add nothink renderer tests (#15554)

Local AiDGX agent

Ollama v0.20.7-rc0 is a release candidate that introduces renderer tests for the `nothink` mode specific to the Gemma 4 model (PR #15554). The `nothink` feature allows Gemma 4 to bypass its default ch

v0.20.8-rc0: Gemma4 on MLX (#15244)

Local AiDGX agent

Ollama v0.20.8-rc0 is a release candidate that introduces MLX support for Google's Gemma 4 model family on Apple Silicon, addressing a prior limitation where Ollama would throw a `Gemma4ForConditional

VAG: Dual-Stream Video-Action Generation for Embodied Data Synthesis

SafetyDGX agent

arXiv:2604.09330v1 Announce Type: cross Abstract: Recent advances in robot foundation models trained on large-scale human teleoperation data have enabled robots to perform increasingly complex real-wo

VAGNet: Vision-based accident anticipation with global features

Model ReleasesDGX agent

arXiv:2604.09305v1 Announce Type: new Abstract: Traffic accidents are a leading cause of fatalities and injuries across the globe. Therefore, the ability to anticipate hazardous situations in advance

Variational Quantum Physics-Informed Neural Networks for Hydrological PDE-Constrained Learning with Inherent Uncertainty Quantification

ResearchDGX agent

arXiv:2604.09374v1 Announce Type: cross Abstract: We propose a Hybrid Quantum-Classical Physics-Informed Neural Network (HQC-PINN) that integrates parameterized variational quantum circuits into the P

Verbalizing LLMs' assumptions to explain and control sycophancy

SafetyDGX agent

arXiv:2604.03058v2 Announce Type: replace-cross Abstract: LLMs can be socially sycophantic, affirming users when they ask questions like 'am I in the wrong?' rather than providing genuine assessment.

VerifAI: A Verifiable Open-Source Search Engine for Biomedical Question Answering

Model ReleasesDGX agent

arXiv:2604.08549v1 Announce Type: cross Abstract: We introduce VerifAI, an open-source expert system for biomedical question answering that integrates retrieval-augmented generation (RAG) with a novel

Very interesting evaluation from the UK’s AI Security Institute of the not yet publicly available Claude Mythos Preview. On the happy side, …

Model ReleasesDGX agent

Very interesting evaluation from the UK’s AI Security Institute of the not yet publicly available Claude Mythos Preview. On the happy side, in its current form, Myth is nowhere near as scary as Tom Fr

Violence is not the answer. But maybe boycotts are?

SafetyDGX agent

Violence is not the answer. But maybe boycotts are? 🚨 NOW: The FBI is RAIDING the home of a 20-year-old man who threw a molotov cocktail at the home of OpenAI CEO Sam Altman Over a DOZEN federal agent

ViSAGE @ NTIRE 2026 Challenge on Video Saliency Prediction

Model ReleasesDGX agent

arXiv:2604.08613v1 Announce Type: new Abstract: In this report, we present our champion solution for the NTIRE 2026 Challenge on Video Saliency Prediction held in conjunction with CVPR 2026. To exploi

Vision Transformers for Preoperative CT-Based Prediction of Histopathologic Chemotherapy Response Score in High-Grade Serous Ovarian Carcinoma

ResearchDGX agent

arXiv:2604.09197v1 Announce Type: cross Abstract: Purpose. High-grade serous ovarian carcinoma (HGSOC) is characterized by pronounced biological and spatial heterogeneity and is frequently diagnosed a

VisionFoundry: Teaching VLMs Visual Perception with Synthetic Images

ResearchDGX agent

arXiv:2604.09531v1 Announce Type: cross Abstract: Vision-language models (VLMs) still struggle with visual perception tasks such as spatial understanding and viewpoint recognition. One plausible contr

VisionLaw: Inferring Interpretable Intrinsic Dynamics from Visual Observations via Bilevel Optimization

ApplicationsDGX agent

arXiv:2508.13792v2 Announce Type: replace Abstract: The intrinsic dynamics of an object governs its physical behavior in the real world, playing a critical role in enabling physically plausible intera

VISOR: Agentic Visual Retrieval-Augmented Generation via Iterative Search and Over-horizon Reasoning

SafetyDGX agent

arXiv:2604.09508v1 Announce Type: cross Abstract: Visual Retrieval-Augmented Generation (VRAG) empowers Vision-Language Models to retrieve and reason over visually rich documents. To tackle complex qu

Visually-Guided Policy Optimization for Multimodal Reasoning

SafetyDGX agent

arXiv:2604.09349v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has significantly advanced the reasoning ability of vision-language models (VLMs). However, the

VL-Calibration: Decoupled Confidence Calibration for Large Vision-Language Models Reasoning

ResearchDGX agent

arXiv:2604.09529v1 Announce Type: cross Abstract: Large Vision Language Models (LVLMs) achieve strong multimodal reasoning but frequently exhibit hallucinations and incorrect responses with high certa

VOLTA: The Surprising Ineffectiveness of Auxiliary Losses for Calibrated Deep Learning

Model ReleasesDGX agent

arXiv:2604.08639v1 Announce Type: cross Abstract: Uncertainty quantification (UQ) is essential for deploying deep learning models in safety critical applications, yet no consensus exists on which UQ m

Voters in Festus, Missouri, ousted all four incumbent council members running for reelection last week, days after the council's approval of a $6B data center (Jeff Tomich/Politico)

IndustryDGX agent

Jeff Tomich / Politico: Voters in Festus, Missouri, ousted all four incumbent council members running for reelection last week, days after the council's approval of a $6B data center — The rout of hal

VSI: Visual Subtitle Integration for Keyframe Selection to enhance Long Video Understanding

ResearchDGX agent

arXiv:2508.06869v4 Announce Type: replace-cross Abstract: Multimodal large language models (MLLMs) demonstrate exceptional performance in vision-language tasks, yet their processing of long videos is

Vultr at HumanX 2026: Shaping AI Strategies for Long-Term Success

Model ReleasesDGX agent

Vultr participated in the HumanX 2026 conference in San Francisco, an event focused on enterprise AI adoption and strategy. The blog post likely highlights Vultr's presence at the conference, showcasi

WAND: Windowed Attention and Knowledge Distillation for Efficient Autoregressive Text-to-Speech Models

ResearchDGX agent

arXiv:2604.08558v1 Announce Type: cross Abstract: Recent decoder-only autoregressive text-to-speech (AR-TTS) models produce high-fidelity speech, but their memory and compute costs scale quadratically

← Previous
1…13491350135113521353…1396
Next →