AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
Human
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
84,561 results
4 Aug 2026

MDTD-ArtIR: Benchmarking Image Editing and Restoration Models for Art Image Restoration under Texture-Overlay Degradations

Model ReleasesDGX agent

arXiv:2608.00736v1 Announce Type: new Abstract: Restoring severely degraded visual media still remains a formidable challenge, as existing methods often hallucinate unnatural textures and contents, st

MDWD: A Street-Level Dataset for Municipal Solid Waste Detection in Dense Urban Environments

Model ReleasesDGX agent

arXiv:2608.00257v1 Announce Type: new Abstract: Automated visual monitoring of urban environments is a growing Computer Vision research area, but municipal solid waste detection remains under-represen

Measuring in-context algorithmic reasoning in language models against an exact Bayes-optimal standard

Model Releases
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2608.01575v1 Announce Type: new Abstract: Whether large language models perform genuine algorithmic reasoning or mere pattern completion is hard to test, because most benchmarks lack a ground tr

Measuring Product Quality Using Images: The CLIP Q-Score and an Application to Real Estate

ResearchDGX agent

arXiv:2608.01544v1 Announce Type: cross Abstract: The CLIP Q-score is a novel, safe, fully reproducible, and computationally efficient method for extracting objective product quality metrics from visu

MedPRESS: A Multi-turn Benchmark for Patient-Pressure-Induced Medical Sycophancy in LLMs

Model ReleasesDGX agent

arXiv:2608.02520v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used for health-related advice. Existing research measures their safety with static questions rather than

MedSAM2-Anatomy: Training-Free Inference-Time Optimization for Musculoskeletal Segmentation

SafetyDGX agent

arXiv:2608.00195v1 Announce Type: cross Abstract: High-resolution 3D segmentation of hip and shoulder anatomy from CT and MRI is essential for surgical planning, yet frozen segmentation models often f

MedTextWeaver: Procedural Knowledge Evolution in Agentic Medical Text Editing

AgentsDGX agent

arXiv:2602.00740v2 Announce Type: replace Abstract: Medical text editing is essential for improving communication among diverse stakeholders in clinical settings. However, adapting LLM agents to this

MedUPS: Towards Diagnostic Assistance in Uncommon Medical Cases with Large Language Models

SafetyDGX agent

arXiv:2608.01012v1 Announce Type: new Abstract: Uncommon and off-guideline cases are difficult for clinical decision support, because physicians must make a series of management decisions under diagno

Meganeura: Portable GPU Training and Inference through Vulkan and Metal

SafetyDGX agent

arXiv:2608.01563v1 Announce Type: new Abstract: Training and deployed inference often cross export, conversion, and platform-specific runtime boundaries. Meganeura asks whether one compact native comp

MemoAct: Atkinson-Shiffrin-Inspired Hierarchical Memory-Augmented Policy for Robotic Manipulation

SafetyDGX agent

arXiv:2603.18494v2 Announce Type: replace Abstract: Memory-augmented robotic policies are essential in handling memory-dependent tasks. However, existing approaches typically rely on simply extending

MemoryForge: Synthesize Lifelong Memory for Human-Like LLM Agents

AgentsDGX agent

arXiv:2608.00007v1 Announce Type: new Abstract: Equipping Large Language Models (LLMs) with human-like personas is crucial for agentic applications, such as role-play and user simulation. Traditional

MemSIF: From Structured Interactions to Dual-Track Fact Memory for LLM Agents

ResearchDGX agent

arXiv:2608.01742v1 Announce Type: cross Abstract: Long-term memory is critical for LLM agents operating over long-horizon interactions. However, several persistent limitations of existing memory syste

Messages, Not Tokens: Grounded Coresets for Faithful VLM Compression

ResearchDGX agent

arXiv:2608.02134v1 Announce Type: new Abstract: Modern vision language models (VLMs) turn high-resolution images into long sequences of visual tokens. Every token traverses the language decoder and pe

MetaRoute-Bench: Evaluating Meta-Decision Policies for Agentic Workflow Routing

Model ReleasesDGX agent

arXiv:2608.00107v1 Announce Type: new Abstract: Agentic systems must repeatedly decide whether to answer directly, decompose a task, invoke a tool, execute code, delegate to a specialist, verify an in

MIDAL: Math Image Descriptions for Accessible Learning

TutorialsDGX agent

arXiv:2608.00868v1 Announce Type: new Abstract: Many open educational resources are lacking in accessibility, especially in-depth image descriptions. In subjects like Science and Mathematics, however,

MIEScore: Human-Aligned Evaluation for Multi-Source Image Editing

Model ReleasesDGX agent

arXiv:2608.02059v1 Announce Type: new Abstract: Recent advances in unified multimodal models have significantly improved text-guided image editing abilities. In particular, models such as Nano-Banana-

Mind the Gap: Zero-Query Jailbreaks via Filter-Generator Discrepancy in Text-to-Image Systems

SafetyDGX agent

arXiv:2608.00973v1 Announce Type: new Abstract: Text-to-image (T2I) systems typically have prompt-level safety filters before the generator to block unsafe requests, yet such systems remain vulnerable

MiniWorld: Democratizing the Training of Video World Models from Scratch

HardwareDGX agent

arXiv:2608.01127v1 Announce Type: new Abstract: Video world models predict future observations conditioned on historical observations and control signals, enabling long-horizon generation through auto

Minute-Scale Training for Microrobot Navigation

Model ReleasesDGX agent

arXiv:2608.00854v1 Announce Type: new Abstract: Microrobots hold significant potential for various applications, where targeted navigation is a basic requirement. Deep reinforcement learning (DRL) has

Mistral releases Shieldstral, a 3B multimodal safety classifier that it says matches models up to 7x its size on text safety, available under Apache 2.0 (Mistral AI Blog)

Model ReleasesDGX agent

Mistral AI Blog: Mistral releases Shieldstral, a 3B multimodal safety classifier that it says matches models up to 7x its size on text safety, available under Apache 2.0 — Every product that ships a m

Mitigating Backdoors via Decoy Shortcuts and Knowledge Decoupling

Model ReleasesDGX agent

arXiv:2608.00732v1 Announce Type: cross Abstract: Backdoor attacks pose a serious threat to deep neural networks, especially when training relies on third-party data, allowing adversaries to inject ma

Mitigating Visual Degradation in MLLMs via Spatial-Spectral Visual Anchor Learning

SafetyDGX agent

arXiv:2608.01635v1 Announce Type: new Abstract: Despite the progress of multimodal large language models (MLLMs), they continue to exhibit deficiencies in visual perception. Following visual instructi

Mitigating Visual Hallucinations in Multimodal Systems through Retrieval-Augmented Reliability-Aware Inference

SafetyDGX agent

arXiv:2606.15782v2 Announce Type: replace-cross Abstract: Multimodal large language models (MLLMs) have demonstrated strong capabilities in vision-language understanding and natural-language response

MixedComplementarityProblems.jl: A Fast, Batched, Open-Source Interior Point Solver for Mixed Complementarity Problems

Model ReleasesDGX agent

arXiv:2608.00959v1 Announce Type: cross Abstract: Mixed complementarity problems (MCPs) arise as the first-order optimality conditions of nonlinear programs and noncooperative games, and provide a nat

MMPhysVideo: Physically Plausible Video Generation Through Joint RGB-Perception Modeling

SafetyDGX agent

arXiv:2604.02817v2 Announce Type: replace Abstract: Despite advancements in generating visually stunning content, video diffusion models (VDMs) often yield physically inconsistent results due to pixel

MoCRA: Mixture of Compositional Rank-1 Atoms for 4K All-in-One Video Restoration

Model ReleasesDGX agent

arXiv:2608.01829v1 Announce Type: new Abstract: Real-world video arrives hazy, rainy, dark, or noisy, and a deployable restorer faces three demands at once: no degradation label, native 4K output, and

Model-Agnostic FDR Control via Group Gaussian Mirror and Permutation SHAP

ApplicationsDGX agent

arXiv:2608.00989v1 Announce Type: cross Abstract: Most FDR-controlled feature selection methods are designed for coordinate-wise hypotheses, where each feature has a single weight or importance score.

Modeling Unknown Nonlocal PDE Systems via Flow Map Learning

TutorialsDGX agent

arXiv:2608.00400v1 Announce Type: new Abstract: Nonlocal partial differential equations arise in many applications but are often difficult to model and learn because of the presence of nonlocal operat

Models as Tools: An Agentic Coordination Framework for Unified Multimodal Visual Tracking

Model ReleasesDGX agent

arXiv:2608.00847v1 Announce Type: new Abstract: Most current visual trackers adopt a matching-based architecture trained exclusively on tracking datasets, whose performance gains depend heavily on the

MonitorVLM-v2: A Deployed Vision-Language Framework for Real-Time Safety Violation Detection

SafetyDGX agent

arXiv:2608.00975v1 Announce Type: new Abstract: Large vision--language models (VLMs) can reason step by step about complex visual scenes, but this open-ended, autoregressive chain-of-thought (CoT) app

MoRAL: Sensor-Grounded BEV Reasoning for Compact VLMs toward Edge-Oriented Autonomous Driving

Model ReleasesDGX agent

arXiv:2608.02449v1 Announce Type: new Abstract: Deploying vision-language models (VLMs) for safety-critical spatial reasoning on resource-constrained autonomous driving platforms requires both compact

Morphology Aware Reversible Semantic Tokenization and Hierarchical Word Composition for Tamil Language Models

Model ReleasesDGX agent

arXiv:2608.01153v1 Announce Type: new Abstract: Statistical subword tokenizers can process arbitrary text, but their units need not align with lexical or grammatical structure. This is especially impo

Motion Beyond Morphology: Bootstrapping Cross-Category Motion Transfer from Abstract Motion Representations

ResearchDGX agent

arXiv:2608.01628v1 Announce Type: new Abstract: Video motion transfer aims to animate a target object using dynamics from a reference video. Existing formulations largely rely on fixed structural corr

Motion Planning for Mobile Manipulators Navigating Doorways via Model Predictive Control

ResearchDGX agent

arXiv:2608.00206v1 Announce Type: new Abstract: Navigating doorways is a fundamental capability for mobile manipulators operating in human environments, requiring coordinated motion between the mobile

Move What Matters: Parameter-Efficient Domain Adaptation via Optimal Transport Flow for Collaborative Perception

Model ReleasesDGX agent

arXiv:2602.11565v5 Announce Type: replace Abstract: Efficient domain adaptation remains a fundamental challenge for deploying multi-agent systems across diverse environments in Vehicle-to-Everything (

Multi-Source Dynamic Graph Learning for Compound-Flood Forecasting in Managed Coastal Systems

SafetyDGX agent

arXiv:2608.01775v1 Announce Type: new Abstract: Compound flooding in managed coastal systems is influenced by hydrological conditions and water-management activity observed across multiple monitoring

Multi-View Unified Camera Fields: Geometry-Shaped Action-Facing Representations for RGB-Only Multi-Camera VLA Policies

ApplicationsDGX agent

arXiv:2608.01826v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have shown strong generalization in robotic manipulation, yet complex contact-rich tasks often benefit from multi-ca

Multiple result sets: How Database Migration Service automates SQL server to PostgreSQL translation

Model ReleasesDGX agent

In the Medium blog post, 'From MARS to SETOF REFCURSOR: Migrating Multi-Result Stored Procedures to PostgreSQL,' we explored the fundamental architectural differences between SQL Server and PostgreSQL

MUSS: Multilevel Subset Selection for Relevance and Diversity

ApplicationsDGX agent

arXiv:2503.11126v4 Announce Type: replace Abstract: The problem of relevant and diverse subset selection has a wide range of applications, including recommender systems and retrieval-augmented generat

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model

SafetyDGX agent

arXiv:2603.14686v2 Announce Type: replace Abstract: Human-Object Interaction (HOI) video reenactment aims to transfer the interaction dynamics of a source video to a novel target object while preservi

Native Multilingual Chain-of-Thought Reasoning in Low-Resource Southeast Asian Languages

SafetyDGX agent

arXiv:2608.00533v1 Announce Type: new Abstract: Large Language Models have achieved substantial progress in reasoning capabilities. Yet in low-resource native settings, many suffer from cross-lingual

Near-Optimal Reinforcement Learning for Constrained Recurrence Objectives

SafetyDGX agent

arXiv:2511.19849v2 Announce Type: replace-cross Abstract: Recurrence objectives, where a target region must be visited infinitely often, are a fundamental class of specifications for Markov decision p

NetDiff: Graph Diffusion with Improved Global Capabilities to Generate and Update Mobile Network Topologies

ResearchDGX agent

arXiv:2410.08238v2 Announce Type: replace-cross Abstract: We introduce NetDiff, a node-conditioned denoising diffusion model that generates directional link topologies and a two-slot transmit/receive

Network Information Enhances Unreliable News Domain Detection

ResearchDGX agent

arXiv:2608.02399v1 Announce Type: cross Abstract: Content-based detection of unreliable news is increasingly difficult, as low-reliability sources mimic credible journalism and generative AI makes fab

Neural Born Series Operator for Biomedical Ultrasound Computed Tomography

ResearchDGX agent

arXiv:2312.15575v2 Announce Type: replace-cross Abstract: Ultrasound Computed Tomography (USCT) provides a radiation-free option for high-resolution clinical imaging. Despite its potential, the comput

Neural Circuit Function Inference with LLMs

ResearchDGX agent

arXiv:2608.00059v1 Announce Type: new Abstract: The success of connectome mapping now shifts the challenge of understanding the nervous system to the interpretation of neural circuits. Here, we devise

Neural operator learning for collision-aware trajectory planning of spacecraft swarms

SafetyDGX agent

arXiv:2608.00320v1 Announce Type: new Abstract: Autonomous spacecraft swarms must plan fuel-efficient, collision-free maneuvers in increasingly congested orbits, yet classical trajectory optimization

Neural Surrogate HMC: On Using Neural Likelihoods for Hamiltonian Monte Carlo in Simulation-Based Inference

ResearchDGX agent

arXiv:2407.20432v3 Announce Type: replace Abstract: Bayesian inference methods such as Markov Chain Monte Carlo (MCMC) typically require repeated computations of the likelihood function, but in some s

New release of LLM adds support for reasoning traces, OpenAI Responses, server-side tools, and smarter logging

Model ReleasesDGX agent

I released LLM 0.32 this morning, the most significant new version of LLM since the initial launch of the project. The new version includes support for visible reasoning traces, server-side provider t

New York Smells: A Large Multimodal Dataset for Olfaction

Model ReleasesDGX agent

arXiv:2511.20544v2 Announce Type: replace Abstract: While olfaction is central to how animals perceive the world, this rich chemical sensory modality remains largely inaccessible to machines. One key

NISF++: Geometrically-grounded implicit representations of 3D+time cardiac function from 2D short- and long-axis MR views

SafetyDGX agent

arXiv:2608.00752v1 Announce Type: new Abstract: Clinical acquisition in cardiac magnetic resonance (CMR) imaging involves obtaining cross-sectional planes of the heart along the radial and longitudina

No One Wins in Nuclear War: A Social Simulation of Military Decision-making

SafetyDGX agent

arXiv:2608.01868v1 Announce Type: cross Abstract: WOPR is a social-simulation environment for studying how organizations make high-stakes decisions, built on a deterministic, replay-validated rules en

Noise-Robust Conditional Flow Matching: Generating Clean Samples from Noisy Datasets

TutorialsDGX agent

arXiv:2608.00064v1 Announce Type: new Abstract: Generative models learn the statistical properties of their training data, so high-quality generation depends on clean and representative datasets. In s

Non-KKT Accumulation in Entropic Mirror Descent

ResearchDGX agent

arXiv:2608.01658v1 Announce Type: cross Abstract: For mirror descent generated by a Legendre kernel, perhaps one of the most basic question in optimization is this: must every accumulation point of a

Nonlinear Laplacians Improve Signed-Directed Graph Learning

ResearchDGX agent

arXiv:2608.00836v1 Announce Type: new Abstract: While signed-directed graphs have been studied using linear Laplacians in the design of graph neural networks, relatively little research has focused on

Nonparametric Distribution Regression Re-calibration

SafetyDGX agent

arXiv:2602.13362v2 Announce Type: replace-cross Abstract: A key challenge in probabilistic regression is ensuring that predictive distributions accurately reflect true empirical uncertainty. Minimizin

Not All EEG Moments Are Equal: Position-Adaptive Time Scheduling for EEG Generation

SafetyDGX agent

arXiv:2608.00048v1 Announce Type: cross Abstract: Electroencephalography (EEG) generation is essential for alleviating data scarcity and enabling large scale neural modeling in brain computer interfac

‘Not healthy’ LLM use is more common than you think

ApplicationsDGX agent

Hank Green, a popular YouTuber and science communicator, said he is stepping back from production amid intense criticism over his use of AI. Green described his AI usage as 'not healthy,' but stressed

Not the Dimension, the Norm: What Matters in Gradient-Free Weight Perturbation of Language Models

Model ReleasesDGX agent

arXiv:2608.01624v1 Announce Type: new Abstract: Adapting a language model to a task no longer requires training all of its weights, and a line of parameter-efficient methods has driven the trainable c

Nova: An End-to-End MLIR Compiler for Deep Learning

Model ReleasesDGX agent

arXiv:2608.00029v1 Announce Type: cross Abstract: The performance of deep learning models at scale relies heavily on how effectively high-level mathematical operations are mapped to underlying physica

← Previous
1…118119120121122…1410
Next →