AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,356
  • Agents7,554
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,165
  • Local Ai4,929
  • Model Releases23,869
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,356
  • Agents7,554
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,165
  • Local Ai4,929
  • Model Releases23,869
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

88,356Total entries
1Added by human
88,355Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,584 results
17 Apr 2026

R3D: Revisiting 3D Policy Learning

SafetyDGX agent

arXiv:2604.15281v1 Announce Type: new Abstract: 3D policy learning promises superior generalization and cross-embodiment transfer, but progress has been hindered by training instabilities and severe o

Rethinking AI Hardware: A Three-Layer Cognitive Architecture for Autonomous Agents

Local AiDGX agent

arXiv:2604.13757v1 Announce Type: new Abstract: The next generation of autonomous AI systems will be constrained not only by model capability, but by how intelligence is structured across heterogeneou

Rethinking Patient Education as Multi-turn Multi-modal Interaction

Model ReleasesDGX agent

arXiv:2604.14656v1 Announce Type: cross Abstract: Most medical multimodal benchmarks focus on static tasks such as image question answering, report generation, and plain-language rewriting. Patient ed

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

S2AM3D: Scale-controllable Part Segmentation of 3D Point Cloud

ResearchDGX agent

arXiv:2512.00995v3 Announce Type: replace Abstract: Part-level point cloud segmentation has recently attracted significant attention in 3D computer vision. Nevertheless, existing research is constrain

Speak, Segment, Track, Navigate: An Interactive System for Video-Guided Skull-Base Surgery

Model ReleasesDGX agent

arXiv:2603.16024v2 Announce Type: replace Abstract: We introduce a speech-guided embodied agent framework for video-guided skull base surgery that dynamically executes perception and image-guidance ta

TableNet A Large-Scale Table Dataset with LLM-Powered Autonomous

AgentsDGX agent

arXiv:2604.13041v1 Announce Type: cross Abstract: Table Structure Recognition (TSR) requires the logical reasoning ability of large language models (LLMs) to handle complex table layouts, but current

The Acoustic Camouflage Phenomenon: Re-evaluating Speech Features for Financial Risk Prediction

ResearchDGX agent

arXiv:2604.14619v1 Announce Type: cross Abstract: In computational paralinguistics, detecting cognitive load and deception from speech signals is a heavily researched domain. Recent efforts have attem

Think in Latent Thoughts: A New Paradigm for Gloss-Free Sign Language Translation

Model ReleasesDGX agent

arXiv:2604.15301v1 Announce Type: new Abstract: Many SLT systems quietly assume that brief chunks of signing map directly to spoken-language words. That assumption breaks down because signers often cr

Time-RA: Towards Time Series Reasoning for Anomaly Diagnosis with LLM Feedback

Model ReleasesDGX agent

arXiv:2507.15066v5 Announce Type: replace Abstract: Time series anomaly detection (TSAD) has traditionally focused on binary classification and often lacks the fine-grained categorization and explanat

TwinOR: Photorealistic Digital Twins of Dynamic Operating Rooms for Embodied AI Research

SafetyDGX agent

arXiv:2511.07412v2 Announce Type: replace Abstract: Developing embodied AI for intelligent surgical systems requires safe, controllable environments for continual learning and evaluation. However, saf

VisRet: Visualization Improves Knowledge-Intensive Text-to-Image Retrieval

Model ReleasesDGX agent

arXiv:2505.20291v4 Announce Type: replace-cross Abstract: Text-to-image retrieval (T2I retrieval) remains challenging because cross-modal embeddings often behave as bags of concepts, underrepresenting

VoxSafeBench: Not Just What Is Said, but Who, How, and Where

SafetyDGX agent

arXiv:2604.14548v1 Announce Type: cross Abstract: As speech language models (SLMs) transition from personal devices into shared, multi-user environments, their responses must account for far more than

What Is the Minimum Architecture for Prolepsis? Early Irrevocable Commitment Across Tasks in Small Transformers

Model ReleasesDGX agent

arXiv:2604.15010v1 Announce Type: cross Abstract: When do transformers commit to a decision, and what prevents them from correcting it? We introduce extbf{prolepsis}: a transformer commits early, task

Which bird does not have wings: Negative-constrained KGQA with Schema-guided Semantic Matching and Self-directed Refinement

ApplicationsDGX agent

arXiv:2604.14749v1 Announce Type: new Abstract: Large language models still struggle with faithfulness and hallucinations despite their remarkable reasoning abilities. In Knowledge Graph Question Answ

WILD-SAM: Phase-Aware Expert Adaptation of SAM for Landslide Detection in Wrapped InSAR Interferograms

Model ReleasesDGX agent

arXiv:2604.14540v1 Announce Type: new Abstract: Detecting slow-moving landslides directly from wrapped Interferometric Synthetic Aperture Radar (InSAR) interferograms is crucial for efficient geohazar

16 Apr 2026

Agent evals are drifting away from production reality. Most benchmarks use clean tasks, well-specified requirements, deterministic metrics, …

Model ReleasesDGX agent

Agent evals are drifting away from production reality. Most benchmarks use clean tasks, well-specified requirements, deterministic metrics, and retrospective curation. Production work is messier, with

Amazon launches its first smart warehouse in Shenzhen, aiming to cut local merchant storage costs by up to 45% as competition with Shein and Temu intensifies (Iris Deng/South China Morning Post)

Model ReleasesDGX agent

Iris Deng / South China Morning Post: Amazon launches its first smart warehouse in Shenzhen, aiming to cut local merchant storage costs by up to 45% as competition with Shein and Temu intensifies — Am

Anthropic says Opus 4.7 hits 80.6% on Document Reasoning — up from 57.1%. But 'reasoning about documents' ≠ 'parsing documents for agents.' …

Model ReleasesDGX agent

Anthropic says Opus 4.7 hits 80.6% on Document Reasoning — up from 57.1%. But 'reasoning about documents' ≠ 'parsing documents for agents.' We ran it on ParseBench. → Charts: 13.5% → 55.8% (+42.3) — h

Artificial intelligence application in lymphoma diagnosis with Vision Transformer using weakly supervised training

ResearchDGX agent

arXiv:2604.13795v1 Announce Type: new Abstract: Vision transformers (ViT) have been shown to allow for more flexible feature detection and can outperform convolutional neural network (CNN) when pre-tr

Autonomous Multi-objective Alloy Design through Simulation-guided Optimization

Model ReleasesDGX agent

arXiv:2507.16005v2 Announce Type: replace-cross Abstract: Alloy discovery is constrained by vast compositional spaces, competing objectives, and prohibitive experimental costs. Although simulations an

Bias at the End of the Score

SafetyDGX agent

arXiv:2604.13305v1 Announce Type: new Abstract: Reward models (RMs) are inherently non-neutral value functions designed and trained to encode specific objectives, such as human preferences or text-ima

Bias-Corrected Adaptive Conformal Inference for Multi-Horizon Time Series Forecasting

SafetyDGX agent

arXiv:2604.13253v1 Announce Type: new Abstract: Adaptive Conformal Inference (ACI) provides distribution-free prediction intervals with asymptotic coverage guarantees for time series under distributio

Binomial Gradient-Based Meta-Learning for Enhanced Meta-Gradient Estimation

ResearchDGX agent

arXiv:2604.13263v1 Announce Type: new Abstract: Meta-learning offers a principled framework leveraging task-invariant priors from related tasks, with which task-specific models can be fine-tuned on do

Can Cross-Layer Transcoders Replace Vision Transformer Activations? An Interpretable Perspective on Vision

ResearchDGX agent

arXiv:2604.13304v1 Announce Type: new Abstract: Understanding the internal activations of Vision Transformers (ViTs) is critical for building interpretable and trustworthy models. While Sparse Autoenc

CANVAS: Continuity-Aware Narratives via Visual Agentic Storyboarding

Model ReleasesDGX agent

arXiv:2604.13452v1 Announce Type: new Abstract: Long-form visual storytelling requires maintaining continuity across shots, including consistent characters, stable environments, and smooth scene trans

CausalDisenSeg: A Causality-Guided Disentanglement Framework with Counterfactual Reasoning for Robust Brain Tumor Segmentation Under Missing Modalities

SafetyDGX agent

arXiv:2604.13409v1 Announce Type: new Abstract: In clinical practice, the robustness of deep learning models for multimodal brain tumor segmentation is severely compromised by incomplete MRI data. Thi

Claude Opus 4.7 is now available in Windsurf 2.0! Anthropic has clearly optimized Claude Opus 4.7 for sustained reasoning over long runs. Ag…

Model ReleasesDGX agent

Claude Opus 4.7 is now available in Windsurf 2.0! Anthropic has clearly optimized Claude Opus 4.7 for sustained reasoning over long runs. Agents stay on track longer without intervention, so engineers

Common to Whom? Regional Cultural Commonsense and LLM Bias in India

Model ReleasesDGX agent

arXiv:2601.15550v3 Announce Type: replace Abstract: Existing cultural commonsense benchmarks treat nations as monolithic, assuming uniform practices within national boundaries. But does cultural commo

Design and Behavior of Sparse Mixture-of-Experts Layers in CNN-based Semantic Segmentation

ResearchDGX agent

arXiv:2604.13761v1 Announce Type: new Abstract: Sparse mixture-of-experts (MoE) layers have been shown to substantially increase model capacity without a proportional increase in computational cost an

Drowsiness-Aware Adaptive Autonomous Braking System based on Deep Reinforcement Learning for Enhanced Road Safety

Model ReleasesDGX agent

arXiv:2604.13878v1 Announce Type: new Abstract: Driver drowsiness significantly impairs the ability to accurately judge safe braking distances and is estimated to contribute to 10%-20% of road acciden

Efficient Multi-View 3D Object Detection by Dynamic Token Selection and Fine-Tuning

Model ReleasesDGX agent

arXiv:2604.13586v1 Announce Type: new Abstract: Existing multi-view three-dimensional (3D) object detection approaches widely adopt large-scale pre-trained vision transformer (ViT)-based foundation mo

For those not seeing the increase, make sure you're using Opus 4.7 with the latest Claude Code

Model ReleasesDGX agent

Claude Code users should ensure they are using Opus 4.7 with the latest updates to experience performance improvements or feature enhancements. The post suggests that users not observing expected incr

From Prediction to Justification: Aligning Sentiment Reasoning with Human Rationale via Reinforcement Learning

ResearchDGX agent

arXiv:2604.13398v1 Announce Type: new Abstract: While Aspect-based Sentiment Analysis (ABSA) systems have achieved high accuracy in identifying sentiment polarities, they often operate as 'black boxes

Frozen Forecasting: A Unified Evaluation

ResearchDGX agent

arXiv:2507.13942v2 Announce Type: replace Abstract: Forecasting future events is a fundamental capability for general-purpose systems that plan or act across different levels of abstraction. Yet, eval

Gemini can now create personalized AI images by digging around in Google Photos

Model ReleasesDGX agent

Gemini now uses user interests and Google Photos to create personalized AI images without requiring long descriptions, allowing users to simply ask for pictures of themselves or family members. This f

HINTBench: Horizon-agent Intrinsic Non-attack Trajectory Benchmark

Model ReleasesDGX agent

arXiv:2604.13954v1 Announce Type: new Abstract: Existing agent-safety evaluation has focused mainly on externally induced risks. Yet agents may still enter unsafe trajectories under benign conditions.

How Automated Reasoning checks in Amazon Bedrock transform generative AI compliance

TutorialsDGX agent

In this post, you'll learn why probabilistic AI validation falls short in regulated industries and how Automated Reasoning checks use formal verification to deliver mathematically proven results. You'

https://x.com/NousResearch/status/2044584517218844775

AgentsDGX agent

Nous Research shared an update or announcement on their official X (Twitter) account, likely relating to their ongoing work in AI model development, fine-tuning, or research releases. Nous Research is

in the grand narrative of Meta x AI, we saw the flop (Llama 4 hurhurhur), and now we’re seeing the turn: - *more* hiring since the soup wars…

Model ReleasesDGX agent

in the grand narrative of Meta x AI, we saw the flop (Llama 4 hurhurhur), and now we’re seeing the turn: - *more* hiring since the soup wars of 2025 - Zuck literally moved in with Alexandr and Nat and

Irregularly Sampled Time Series Interpolation for Binary Evolution Simulations Using Dynamic Time Warping

SafetyDGX agent

arXiv:2604.13604v1 Announce Type: cross Abstract: Binary stellar evolution simulations are computationally expensive. Stellar population synthesis relies on these detailed evolution models at a fundam

Jump-Start Reinforcement Learning with Vision-Language-Action Regularization

SafetyDGX agent

arXiv:2604.13733v1 Announce Type: new Abstract: Reinforcement learning (RL) enables high-frequency, closed-loop control for robotic manipulation, but scaling to long-horizon tasks with sparse or imper

Memp: Exploring Agent Procedural Memory

AgentsDGX agent

arXiv:2508.06433v4 Announce Type: replace Abstract: Large Language Models (LLMs) based agents excel at diverse tasks, yet they suffer from brittle procedural memory that is manually engineered or enta

Merry Claude-mas! Opus 4.7 in Claude Code is a monster. Very happy camper! https://www.anthropic.com/news/claude-opus-4-7

Model ReleasesDGX agent

This post celebrates Claude Opus 4.7's capabilities within Claude Code, expressing enthusiasm about its performance and describing it as exceptionally powerful. The entry references an Anthropic annou

MixAtlas: Uncertainty-aware Data Mixture Optimization for Multimodal LLM Midtraining

ResearchDGX agent

This paper was accepted at the Workshop on Navigating and Addressing Data Problems for Foundation Models (NADPFM) at ICLR 2026. Principled domain reweighting can substantially improve sample efficienc

Motif-Video-2B

Local AiDGX agent

Motif-Video-2B is a 2-billion-parameter open-source video generation diffusion transformer released by Motif Technologies in April 2026, capable of both text-to-video and image-to-video generation fro

Mozilla launches Thunderbolt AI client with focus on self-hosted infrastructure

Model ReleasesDGX agent

Thunderbolt is a new open-source AI client from Mozilla-owned MZLA Technologies aimed at enterprises who want to run self-hosted chatbots on their own infrastructure. The platform allows organizations

🎬 Ollama Gemma Day Recap: SGLang at the Ollama Gemma 4 Party in Palo Alto 🍾 Last night, @ollama hosted a packed Gemma Day at the Palo Alto…

Model ReleasesDGX agent

🎬 Ollama Gemma Day Recap: SGLang at the Ollama Gemma 4 Party in Palo Alto 🍾 Last night, @ollama hosted a packed Gemma Day at the Palo Alto office alongside the @GoogleDeepMind Gemma team. SGLang was i

Optimizing Earth Observation Satellite Schedules under Unknown Operational Constraints: An Active Constraint Acquisition Approach

TutorialsDGX agent

arXiv:2604.13283v1 Announce Type: cross Abstract: Earth Observation (EO) satellite scheduling (deciding which imaging tasks to perform and when) is a well-studied combinatorial optimization problem. E

Opus 4.7 is in Claude Code today. It's more agentic, more precise, and a lot better at long-running work. It carries context across sessions…

Model ReleasesDGX agent

Opus 4.7 is in Claude Code today. It's more agentic, more precise, and a lot better at long-running work. It carries context across sessions and handles ambiguity much better. Introducing Claude Opus

Pareto-Optimal Offline Reinforcement Learning via Smooth Tchebysheff Scalarization

SafetyDGX agent

arXiv:2604.13175v1 Announce Type: new Abstract: Large language models can be aligned with human preferences through offline reinforcement learning (RL) on small labeled datasets. While single-objectiv

PartNerFace: Part-based Neural Radiance Fields for Animatable Facial Avatar Reconstruction

TutorialsDGX agent

arXiv:2604.13918v1 Announce Type: new Abstract: We present PartNerFace, a part-based neural radiance fields approach, for reconstructing animatable facial avatar from monocular RGB videos. Existing so

PatchPoison: Poisoning Multi-View Datasets to Degrade 3D Reconstruction

Model ReleasesDGX agent

arXiv:2604.13153v1 Announce Type: new Abstract: 3D Gaussian Splatting (3DGS) has recently enabled highly photorealistic 3D reconstruction from casually captured multi-view images. However, this access

Quantum Machine Learning for Colorectal Cancer Data: Anastomotic Leak Classification and Risk Factors

ResearchDGX agent

arXiv:2604.13951v1 Announce Type: new Abstract: This study evaluates colorectal risk factors and compares classical models against Quantum Neural Networks (QNNs) for anastomotic leak prediction. Analy

RANDPOL: Parameter-Efficient End-to-End Quadruped Locomotion via Randomized Policy Learning

Model ReleasesDGX agent

arXiv:2505.19054v2 Announce Type: replace Abstract: Modern learning-based locomotion controllers typically rely on fully trainable deep neural networks with a large number of parameters. This paper st

Reconstruction of a 3D wireframe from a single line drawing via generative depth estimation

ResearchDGX agent

arXiv:2604.13549v1 Announce Type: new Abstract: The conversion of 2D freehand sketches into 3D models remains a pivotal challenge in computer vision, bridging the gap between human creativity and digi

Rethinking Image-to-3D Generation with Sparse Queries: Efficiency, Capacity, and Input-View Bias

Local AiDGX agent

arXiv:2604.13905v1 Announce Type: new Abstract: We present SparseGen, a novel framework for efficient image-to-3D generation, which exhibits low input-view bias while being significantly faster. Unlik

SemAttNet: Towards Attention-based Semantic Aware Guided Depth Completion

Model ReleasesDGX agent

arXiv:2204.13635v2 Announce Type: replace Abstract: Depth completion involves recovering a dense depth map from a sparse map and an RGB image. Recent approaches focus on utilizing color images as guid

SHARe-KAN: Post-Training Vector Quantization for Cache-Resident KAN Inference

Model ReleasesDGX agent

arXiv:2512.15742v2 Announce Type: replace Abstract: Pre-trained Vision Kolmogorov-Arnold Networks (KANs) store a dense B-spline grid on every edge, inflating prediction-head parameter counts by more t

Shocking result on my pelican benchmark this morning, I got a better pelican from a 21GB local Qwen3.6-35B-A3B running on my laptop than I d…

Model ReleasesDGX agent

Shocking result on my pelican benchmark this morning, I got a better pelican from a 21GB local Qwen3.6-35B-A3B running on my laptop than I did from the new Opus 4.7! Qwen on the left, Opus on the righ

SocialMirror: Reconstructing 3D Human Interaction Behaviors from Monocular Videos with Semantic and Geometric Guidance

Model ReleasesDGX agent

arXiv:2604.13581v1 Announce Type: new Abstract: Accurately reconstructing human behavior in close-interaction scenarios is crucial for enabling realistic virtual interactions in augmented reality, pre

← Previous
1…581582583584585…1060
Next →