AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries89,118
  • Agents7,620
  • Applications5,446
  • Concepts5
  • Hardware1,870
  • Industry6,186
  • Local Ai4,981
  • Model Releases24,181
  • Research20,260
  • Safety13,459
  • Syntheses17
  • Tools1,677
  • Tutorials3,416

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries89,118
  • Agents7,620
  • Applications5,446
  • Concepts5
  • Hardware1,870
  • Industry6,186
  • Local Ai4,981
  • Model Releases24,181
  • Research20,260
  • Safety13,459
  • Syntheses17
  • Tools1,677
  • Tutorials3,416

Source
HumanDGX agent

89,118Total entries
1Added by human
89,117Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
64,222 results
27 Apr 2026

Time-Localized Parametric Decomposition of Respiratory Airflow for Sub-Breath Analysis

Model ReleasesDGX agent

arXiv:2604.22695v1 Announce Type: cross Abstract: Respiratory airflow signals provide critical insight into breathing mechanics, yet conventional analysis methods remain limited in their ability to ch

Useful nonrobust features are ubiquitous in biomedical images

TutorialsDGX agent

arXiv:2604.22579v1 Announce Type: cross Abstract: We study whether deep networks for medical imaging learn useful nonrobust features - predictive input patterns that are not human interpretable and hi

26 Apr 2026

HealthBench Professional, our new evaluation of clinician chat tasks, is now available on HuggingFace for easy access. Each example was writ…

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
IndustryDGX agent

HealthBench Professional, our new evaluation of clinician chat tasks, is now available on HuggingFace for easy access. Each example was written, reviewed, and adjudicated by three or more physicians.

NEW paper from Alibaba. A 30B MoE with only 3B active params matches Qwen3-235B on real tool-use workloads. AgenticQwen-30B-A3B: 50.2 averag…

Model ReleasesDGX agent

NEW paper from Alibaba. A 30B MoE with only 3B active params matches Qwen3-235B on real tool-use workloads. AgenticQwen-30B-A3B: 50.2 average on TAU-2 + BFCL-V4 Multi-Turn. AgenticQwen-8B: 47.4. Both

Why Ollama Cloud doesn't have DeepSeek V4 Pro and Qwen3.6?

Model ReleasesDGX agent

Ollama Cloud has made DeepSeek-V4-Flash available , but the Reddit discussion likely addresses why the more powerful DeepSeek-V4-Pro—with 1.6T total parameters offering performance rivaling top closed

25 Apr 2026

Comprehensive Multilingual Text Rendering: Better glyph accuracy, more consistent typography, and cleaner layouts even in complex compositio…

Model ReleasesDGX agent

Comprehensive Multilingual Text Rendering: Better glyph accuracy, more consistent typography, and cleaner layouts even in complex compositions. It handles mixed-language scenarios more gracefully as w

24 Apr 2026

A Green-Integral-Constrained Neural Solver with Stochastic Physics-Informed Regularization

Model ReleasesDGX agent

arXiv:2604.21411v1 Announce Type: new Abstract: Standard physics-informed neural networks (PINNs) struggle to simulate highly oscillatory Helmholtz solutions in heterogeneous media because pointwise m

A Metamorphic Testing Approach to Diagnosing Memorization in LLM-Based Program Repair

Model ReleasesDGX agent

arXiv:2604.21579v1 Announce Type: cross Abstract: LLM-based automated program repair (APR) techniques have shown promising results in reducing debugging costs. However, prior results can be affected b

A Replicable Robotics Awareness Method Using LLM-Enabled Robotics Interaction: Evidence from a Corporate Challenge

ResearchDGX agent

arXiv:2604.21377v1 Announce Type: new Abstract: Large language models are increasingly being explored as interfaces between humans and robotic systems, yet there remains limited evidence on how such t

Adaptive Defense Orchestration for RAG: A Sentinel-Strategist Architecture against Multi-Vector Attacks

Model ReleasesDGX agent

arXiv:2604.20932v1 Announce Type: cross Abstract: Retrieval-augmented generation (RAG) systems are increasingly deployed in sensitive domains such as healthcare and law, where they rely on private, do

Addressing divergent representations from causal interventions on neural networks

ResearchDGX agent

arXiv:2511.04638v5 Announce Type: replace-cross Abstract: A common approach to mechanistic interpretability is to causally manipulate model representations via targeted interventions in order to under

An update on recent Claude Code quality reports

Model ReleasesDGX agent

An update on recent Claude Code quality reports It turns out the high volume of complaints that Claude Code was providing worse quality results over the past two months was grounded in real problems.

ATOM: A Pretrained Neural Operator for Multitask Molecular Dynamics

ResearchDGX agent

arXiv:2510.05482v2 Announce Type: replace Abstract: Molecular dynamics (MD) simulations underpin modern computational drug discovery, materials science, and biochemistry. Recent machine learning model

Bridging the Training-Deployment Gap: Gated Encoding and Multi-Scale Refinement for Efficient Quantization-Aware Image Enhancement

ResearchDGX agent

arXiv:2604.21743v1 Announce Type: new Abstract: Image enhancement models for mobile devices often struggle to balance high output quality with the fast processing speeds required by mobile hardware. W

Causal Disentanglement for Full-Reference Image Quality Assessment

ResearchDGX agent

arXiv:2604.21654v1 Announce Type: cross Abstract: Existing deep network-based full-reference image quality assessment (FR-IQA) models typically work by performing pairwise comparisons of deep features

Clinically Interpretable Sepsis Early Warning via LLM-Guided Simulation of Temporal Physiological Dynamics

AgentsDGX agent

arXiv:2604.20924v1 Announce Type: new Abstract: Timely and interpretable early warning of sepsis remains a major clinical challenge due to the complex temporal dynamics of physiological deterioration.

Crystal: Characterizing Relative Impact of Scholarly Publications

SafetyDGX agent

arXiv:2603.26791v2 Announce Type: replace-cross Abstract: Assessing a cited paper's impact is typically done by analyzing its citation context in isolation within the citing paper. While this focuses

Demystifying Action Space Design for Robotic Manipulation Policies

SafetyDGX agent

arXiv:2602.23408v2 Announce Type: replace-cross Abstract: The specification of the action space plays a pivotal role in imitation-based robotic manipulation policy learning, fundamentally shaping the

DiffNR: Diffusion-Enhanced Neural Representation Optimization for Sparse-View 3D Tomographic Reconstruction

ResearchDGX agent

arXiv:2604.21518v1 Announce Type: cross Abstract: Neural representations (NRs), such as neural fields and 3D Gaussians, effectively model volumetric data in computed tomography (CT) but suffer from se

Directional Confusions Reveal Divergent Inductive Biases Through Rate-Distortion Geometry in Human and Machine Vision

SafetyDGX agent

arXiv:2604.21909v1 Announce Type: new Abstract: Humans and modern vision models can reach similar classification accuracy while making systematically different kinds of mistakes - differing not in how

Discriminative-Generative Synergy for Occlusion Robust 3D Human Mesh Recovery

ApplicationsDGX agent

arXiv:2604.21712v1 Announce Type: new Abstract: 3D human mesh recovery from monocular RGB images aims to estimate anatomically plausible 3D human models for downstream applications, but remains challe

ERA: Evidence-based Reliability Alignment for Honest Retrieval-Augmented Generation

SafetyDGX agent

arXiv:2604.20854v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) grounds language models in factual evidence but introduces critical challenges regarding knowledge conflicts betw

Escaping the Agreement Trap: Defensibility Signals for Evaluating Rule-Governed AI

SafetyDGX agent

arXiv:2604.20972v1 Announce Type: new Abstract: Content moderation systems are typically evaluated by measuring agreement with human labels. In rule-governed environments this assumption fails: multip

Evaluating AI Meeting Summaries with a Reusable Cross-Domain Pipeline

Model ReleasesDGX agent

arXiv:2604.21345v1 Announce Type: new Abstract: We present a reusable evaluation pipeline for generative AI applications, instantiated for AI meeting summaries and released with a public artifact pack

Expanding the extreme-k dielectric materials space through physics-validated generative reasoning

ResearchDGX agent

arXiv:2604.21068v1 Announce Type: cross Abstract: The most technologically consequential materials are often the rarest: they occupy narrow regions of chemical space, obey competing physical constrain

Federated Learning for Surgical Vision in Appendicitis Classification: Results of the FedSurg EndoVis 2024 Challenge

Model ReleasesDGX agent

arXiv:2510.04772v2 Announce Type: replace-cross Abstract: Developing generalizable surgical AI requires multi-institutional data, yet patient privacy constraints preclude direct data sharing, making F

Flow Matching for Conditional MRI-CT and CBCT-CT Image Synthesis

Model ReleasesDGX agent

arXiv:2510.04823v2 Announce Type: replace Abstract: Generating synthetic CT (sCT) from MRI or CBCT plays a crucial role in enabling MRI-only and CBCT-based adaptive radiotherapy, improving treatment p

For the last 72 hours since ml-intern launched we have had over 500+ autonomous AI research projects running on the Space at all times. Some…

Model ReleasesDGX agent

For the last 72 hours since ml-intern launched we have had over 500+ autonomous AI research projects running on the Space at all times. Some insane ones I saw: 1. A new AI paradigm from scratch — tryi

Forget, Then Recall: Learnable Compression and Selective Unfolding via Gist Sparse Attention

ResearchDGX agent

arXiv:2604.20920v1 Announce Type: new Abstract: Scaling large language models to long contexts is challenging due to the quadratic computational cost of full attention. Mitigation approaches include K

From Tokens to Concepts: Leveraging SAE for SPLADE

ResearchDGX agent

arXiv:2604.21511v1 Announce Type: cross Abstract: Learned Sparse IR models, such as SPLADE, offer an excellent efficiency-effectiveness tradeoff. However, they rely on the underlying backbone vocabula

Frozen LLMs as Map-Aware Spatio-Temporal Reasoners for Vehicle Trajectory Prediction

Local AiDGX agent

arXiv:2604.21479v1 Announce Type: new Abstract: Large language models (LLMs) have recently demonstrated strong reasoning capabilities and attracted increasing research attention in the field of autono

gpt-5.5 unlocks a new level of possibility:

Model ReleasesDGX agent

gpt-5.5 unlocks a new level of possibility: GPT-5.5 is now available in Devin as an Agent Preview! GPT-5.5 has set a new bar for what's possible with Devin. It runs longer and more autonomously than a

Grounding Machine Creativity in Game Design Knowledge Representations: Empirical Probing of LLM-Based Executable Synthesis of Goal Playable Patterns under Structural Constraints

Model ReleasesDGX agent

arXiv:2603.07101v3 Announce Type: replace Abstract: Creatively translating complex gameplay ideas into executable artifacts (e.g., games as Unity projects and code) remains a central challenge in comp

Hierarchical Policy Optimization for Simultaneous Translation of Unbounded Speech

SafetyDGX agent

arXiv:2604.21045v1 Announce Type: new Abstract: Simultaneous speech translation (SST) generates translations while receiving partial speech input. Recent advances show that large language models (LLMs

How English Print Media Frames Human-Elephant Conflicts in India

SafetyDGX agent

arXiv:2604.21496v1 Announce Type: new Abstract: Human-elephant conflict (HEC) is rising across India as habitat loss and expanding human settlements force elephants into closer contact with people. Wh

Identifying Bias in Machine-generated Text Detection

SafetyDGX agent

arXiv:2512.09292v2 Announce Type: replace-cross Abstract: The meteoric rise in text generation capability has been accompanied by parallel growth in interest in machine-generated text detection: the c

Kernel-Smith: A Unified Recipe for Evolutionary Kernel Optimization

Model ReleasesDGX agent

arXiv:2603.28342v2 Announce Type: replace Abstract: We present Kernel-Smith, a framework for high-performance GPU kernel and operator generation that combines a stable evaluation-driven evolutionary a

KinetiDiff: Docking-Guided Diffusion for De Novo ACVR1 Inhibitor Design in Fibrodysplasia Ossificans Progressiva

ResearchDGX agent

arXiv:2604.20886v1 Announce Type: cross Abstract: We present KinetiDiff, a structure-based framework for de novo kinase inhibitor design that integrates a Geometry-Complete Diffusion Model with real-t

Learning Long-Term Motion Embeddings for Efficient Kinematics Generation

ResearchDGX agent

Understanding and predicting motion is a fundamental component of visual intelligence. Although modern video models exhibit strong comprehension of scene dynamics, exploring multiple possible futures

Modulating Cross-Modal Convergence with Single-Stimulus, Intra-Modal Dispersion

SafetyDGX agent

arXiv:2604.21836v1 Announce Type: cross Abstract: Neural networks exhibit a remarkable degree of representational convergence across diverse architectures, training objectives, and even data modalitie

Nonlinear Causal Discovery through a Sequential Edge Orientation Approach

ApplicationsDGX agent

arXiv:2506.05590v3 Announce Type: replace-cross Abstract: Recent advances have established the identifiability of a directed acyclic graph (DAG) under additive noise models (ANMs), spurring the develo

On Reasoning Behind Next Occupation Recommendation

ApplicationsDGX agent

arXiv:2604.21204v1 Announce Type: cross Abstract: In this work, we develop a novel reasoning approach to enhance the performance of large language models (LLMs) in future occupation prediction. In thi

PanGuide3D: Cohort-Robust Pancreas Tumor Segmentation via Probabilistic Pancreas Conditioning and a Transformer Bottleneck

ResearchDGX agent

arXiv:2604.20981v1 Announce Type: cross Abstract: Pancreatic tumor segmentation in contrast-enhanced computed tomography (CT) is clinically important yet technically challenging: lesions are often sma

Promoting Simple Agents: Ensemble Methods for Event-Log Prediction

ApplicationsDGX agent

arXiv:2604.21629v1 Announce Type: cross Abstract: We compare lightweight automata-based models (n-grams) with neural architectures (LSTM, Transformer) for next-activity prediction in streaming event l

Really impressed by how smooth switching most of my coding tasks to Codex (GPT-5.5) from Claude Code (Opus 4.7) has been. I thought it was g…

Model ReleasesDGX agent

Really impressed by how smooth switching most of my coding tasks to Codex (GPT-5.5) from Claude Code (Opus 4.7) has been. I thought it was going to be more difficult and that I would be 'fighting' wit

Sakana AI entered the commercial API market with the beta launch of 'Sakana Fugu,' a sophisticated multi-agent orchestration system previous…

AgentsDGX agent

Sakana AI entered the commercial API market with the beta launch of 'Sakana Fugu,' a sophisticated multi-agent orchestration system previously utilized as their internal secret weapon. Fugu fundamenta

Sapiens2

TutorialsDGX agent

arXiv:2604.21681v1 Announce Type: new Abstract: We present Sapiens2, a model family of high-resolution transformers for human-centric vision focused on generalization, versatility, and high-fidelity o

SatSAM2: Motion-Constrained Video Object Tracking in Satellite Imagery using Promptable SAM2 and Kalman Priors

Model ReleasesDGX agent

arXiv:2511.18264v3 Announce Type: replace Abstract: Existing satellite video tracking methods often struggle with generalization, requiring scenario-specific training to achieve satisfactory performan

SparKV: Overhead-Aware KV Cache Loading for Efficient On-Device LLM Inference

Local AiDGX agent

arXiv:2604.21231v1 Announce Type: cross Abstract: Efficient inference for on-device Large Language Models (LLMs) remains challenging due to limited hardware resources and the high cost of the prefill

Spatial Metaphors for LLM Memory: A Critical Analysis of the MemPalace Architecture

Model ReleasesDGX agent

arXiv:2604.21284v1 Announce Type: new Abstract: MemPalace is an open-source AI memory system that applies the ancient method of loci (memory palace) spatial metaphor to organize long-term memory for l

StyleID: A Perception-Aware Dataset and Metric for Stylization-Agnostic Facial Identity Recognition

Model ReleasesDGX agent

arXiv:2604.21689v1 Announce Type: cross Abstract: Creative face stylization aims to render portraits in diverse visual idioms such as cartoons, sketches, and paintings while retaining recognizable ide

Temporal Taskification in Streaming Continual Learning: A Source of Evaluation Instability

Model ReleasesDGX agent

arXiv:2604.21930v1 Announce Type: new Abstract: Streaming Continual Learning (CL) typically converts a continuous stream into a sequence of discrete tasks through temporal partitioning. We argue that

Transformer-Progressive Mamba Network for Lightweight Image Super-Resolution

ResearchDGX agent

arXiv:2511.03232v2 Announce Type: replace Abstract: Recently, Mamba-based super-resolution (SR) methods have demonstrated the ability to capture global receptive fields with linear complexity, address

When Bigger Isn't Better: A Comprehensive Fairness Evaluation of Political Bias in Multi-News Summarisation

SafetyDGX agent

arXiv:2604.21309v1 Announce Type: new Abstract: Multi-document news summarisation systems are increasingly adopted for their convenience in processing vast daily news content, making fairness across d

WildSplatter: Feed-forward 3D Gaussian Splatting with Appearance Control from Unconstrained Images

ApplicationsDGX agent

arXiv:2604.21182v1 Announce Type: new Abstract: We propose WildSplatter, a feed-forward 3D Gaussian Splatting (3DGS) model for unconstrained images with unknown camera parameters and varying lighting

23 Apr 2026

5B tokens in ml-intern in 48h 😅😅😅

IndustryDGX agent

5B tokens in ml-intern in 48h 😅😅😅 we burned through 5 billion tokens in 48h. turns out giving everyone unlimited access to the most expensive model on the planet is not a great business strategy 🙈 So

Adaptive Multi-task Learning for Multi-sector Portfolio Optimization

ResearchDGX agent

arXiv:2507.16433v2 Announce Type: replace-cross Abstract: Accurate transfer of information across multiple sectors to enhance model estimation is both significant and challenging in multi-sector portf

[AINews] Tasteful Tokenmaxxing

ToolsDGX agent

This episode of AINews from Latent Space likely discusses strategies for optimizing token usage and efficiency in AI models, focusing on practical approaches to maximizing computational value while ma

Apple fixes a bug that stored notifications for deleted messages on iPhone and iPad, following a report that police used it to extract deleted Signal messages (Lorenzo Franceschi-Bicchierai/TechCrunch)

Model ReleasesDGX agent

Lorenzo Franceschi-Bicchierai / TechCrunch: Apple fixes a bug that stored notifications for deleted messages on iPhone and iPad, following a report that police used it to extract deleted Signal messag

Benchmarking ResNet for Short-Term Hypoglycemia Classification with DiaData

Model ReleasesDGX agent

arXiv:2511.02849v2 Announce Type: replace-cross Abstract: Individualized therapy is driven forward by medical data analysis, which provides insight into the patient's context. In particular, for Type

← Previous
1…489490491492493…1071
Next →