AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
84,532 results
16 Apr 2026

Towards Patient-Specific Deformable Registration in Laparoscopic Surgery

ResearchDGX agent

arXiv:2604.13186v1 Announce Type: new Abstract: Unsafe surgical care is a critical health concern, often linked to limitations in surgeon experience, skills, and situational awareness. Integrating pat

Towards Successful Implementation of Automated Raveling Detection: Effects of Training Data Size, Illumination Difference, and Spatial Shift

Model ReleasesDGX agent

arXiv:2604.13322v1 Announce Type: new Abstract: Raveling, the loss of aggregates, is a major form of asphalt pavement surface distress, especially on highways. While research has shown that machine le

Towards Unconstrained Human-Object Interaction

ResearchDGX agent

arXiv:2604.14069v1 Announce Type: new Abstract: Human-Object Interaction (HOI) detection is a longstanding computer vision problem concerned with predicting the interaction between humans and objects.

Content type
AllBlogX PostPaperYouTubeRedditGitHub

Training and Finetuning Multimodal Embedding & Reranker Models with Sentence Transformers

ToolsDGX agent

This guide covers how to train and fine-tune multimodal embedding and reranker models using the Sentence Transformers library, enabling systems to work with both text and image data simultaneously. It

Training-Free Semantic Multi-Object Tracking with Vision-Language Models

ResearchDGX agent

arXiv:2604.14074v1 Announce Type: new Abstract: Semantic Multi-Object Tracking (SMOT) extends multi-object tracking with semantic outputs such as video summaries, instance-level captions, and interact

Training-Free Test-Time Contrastive Learning for Large Language Models

AgentsDGX agent

arXiv:2604.13552v1 Announce Type: new Abstract: Large language models (LLMs) demonstrate strong reasoning capabilities, but their performance often degrades under distribution shift. Existing test-tim

Transcriptomic Models for Immunotherapy Response Prediction Show Limited Cross-cohort Generalisability

Model ReleasesDGX agent

arXiv:2604.05478v2 Announce Type: replace-cross Abstract: Immune checkpoint inhibitors (ICIs) have transformed cancer therapy; yet substantial proportion of patients exhibit intrinsic or acquired resi

Transform retail with AWS generative AI services

IndustryDGX agent

Online retailers face a persistent challenge: shoppers struggle to determine the fit and look when ordering online, leading to increased returns and decreased purchase confidence. The cost? Lost reven

Treating enterprise AI as an operating layer

Model ReleasesDGX agent

There’s a fault line running through enterprise AI, and it’s not the one getting the most attention. The public conversation still tracks foundation models and benchmarks—GPT versus Gemini, reasoning

TREX: Automating LLM Fine-tuning via Agent-Driven Tree-based Exploration

Model ReleasesDGX agent

arXiv:2604.14116v1 Announce Type: cross Abstract: While Large Language Models (LLMs) have empowered AI research agents to perform isolated scientific tasks, automating complex, real-world workflows, s

TRIM: Hybrid Inference via Targeted Stepwise Routing in Multi-Step Reasoning Tasks

SafetyDGX agent

arXiv:2601.10245v2 Announce Type: replace-cross Abstract: Multi-step reasoning tasks like mathematical problem solving are vulnerable to cascading failures, where a single incorrect step leads to comp

TSMC CEO C.C. Wei says TSMC checked with customers about AI demand and was reassured that it was still strong amid the Iran war, as it raises revenue forecasts (Wall Street Journal)

IndustryDGX agent

Wall Street Journal: TSMC CEO C.C. Wei says TSMC checked with customers about AI demand and was reassured that it was still strong amid the Iran war, as it raises revenue forecasts — Taiwan company ex

TSMC reports Q1 revenue up 35.1% YoY to ~35B, net income up 58.3% YoY to ~18B, both above est., and says 7nm or smaller chips were ~74% of its wafer revenue (Dylan Butts/CNBC)

ApplicationsDGX agent

Dylan Butts / CNBC: TSMC reports Q1 revenue up 35.1% YoY to ~35B, net income up 58.3% YoY to ~18B, both above est., and says 7nm or smaller chips were ~74% of its wafer revenue — Taiwan Semiconductor

Turns out chatgpt plus can run a whole openclaw agent team in the background and mine has been running 3 of them for months

AgentsDGX agent

A Reddit user on r/ChatGPT discovered that a ChatGPT Plus subscription can be used to authenticate and power multiple OpenClaw agent instances simultaneously — in their case, three agents running cont

Two AIs walk into a bar...

Local AiDGX agent

A Reddit post from the r/ollama community titled 'Two AIs walk into a bar...' likely showcases a humorous or experimental multi-agent conversation in which two locally-run AI models (via Ollama) inter

Two Pathways to Truthfulness: On the Intrinsic Encoding of LLM Hallucinations

ResearchDGX agent

arXiv:2601.07422v2 Announce Type: replace Abstract: Despite their impressive capabilities, large language models (LLMs) frequently generate hallucinations. Previous work shows that their internal stat

Two-Stage Regularization-Based Structured Pruning for LLMs

Model ReleasesDGX agent

arXiv:2505.18232v3 Announce Type: replace-cross Abstract: The deployment of large language models (LLMs) is largely hindered by their large number of parameters. Structural pruning has emerged as a pr

UHR-BAT: Budget-Aware Token Compression Vision-Language model for Ultra-High-Resolution Remote Sensing

ResearchDGX agent

arXiv:2604.13565v1 Announce Type: new Abstract: Ultra-high-resolution (UHR) remote sensing imagery couples kilometer-scale context with query-critical evidence that may occupy only a few pixels. Such

UI-Copilot: Advancing Long-Horizon GUI Automation via Tool-Integrated Policy Optimization

Model ReleasesDGX agent

arXiv:2604.13822v1 Announce Type: new Abstract: MLLM-based GUI agents have demonstrated strong capabilities in complex user interface interaction tasks. However, long-horizon scenarios remain challeng

UI-Zoomer: Uncertainty-Driven Adaptive Zoom-In for GUI Grounding

Local AiDGX agent

arXiv:2604.14113v1 Announce Type: cross Abstract: GUI grounding, which localizes interface elements from screenshots given natural language queries, remains challenging for small icons and dense layou

UMI-3D: Extending Universal Manipulation Interface from Vision-Limited to 3D Spatial Perception

SafetyDGX agent

arXiv:2604.14089v1 Announce Type: new Abstract: We present UMI-3D, a multimodal extension of the Universal Manipulation Interface (UMI) for robust and scalable data collection in embodied manipulation

UNBOX: Unveiling Black-box visual models with Natural-language

SafetyDGX agent

arXiv:2603.08639v2 Announce Type: replace Abstract: Ensuring trustworthiness in open-world visual recognition requires models that are interpretable, fair, and robust to distribution shifts. Yet moder

UniBlendNet: Unified Global, Multi-Scale, and Region-Adaptive Modeling for Ambient Lighting Normalization

Model ReleasesDGX agent

arXiv:2604.13383v1 Announce Type: new Abstract: Ambient Lighting Normalization (ALN) aims to restore images degraded by complex, spatially varying illumination conditions. Existing methods, such as IF

UniGeoSeg: Towards Unified Open-World Segmentation for Geospatial Scenes

Model ReleasesDGX agent

arXiv:2511.23332v2 Announce Type: replace Abstract: Instruction-driven segmentation in remote sensing generates masks from guidance, offering great potential for accessible and generalizable applicati

Universality of Gaussian-Mixture Reverse Kernels in Conditional Diffusion

ResearchDGX agent

arXiv:2604.13470v1 Announce Type: new Abstract: We prove that conditional diffusion models whose reverse kernels are finite Gaussian mixtures with ReLU-network logits can approximate suitably regular

Unleashing Implicit Rewards: Prefix-Value Learning for Distribution-Level Optimization

Local AiDGX agent

arXiv:2604.13197v1 Announce Type: new Abstract: Process reward models (PRMs) provide fine-grained reward signals along the reasoning process, but training reliable PRMs often requires step annotations

UNRIO: Uncertainty-Aware Velocity Learning for Radar-Inertial Odometry

Model ReleasesDGX agent

arXiv:2604.13584v1 Announce Type: new Abstract: We present UNRIO, an uncertainty-aware radar-inertial odometry system that estimates ego-velocity directly from raw mmWave radar IQ signals rather than

Unsupervised Anomaly Detection in Process-Complex Industrial Time Series: A Real-World Case Study

Model ReleasesDGX agent

arXiv:2604.13928v1 Announce Type: new Abstract: Industrial time-series data from real production environments exhibits substantially higher complexity than commonly used benchmark datasets, primarily

Unsupervised domain transfer: Overcoming signal degradation in sleep monitoring by increasing scoring realism

TutorialsDGX agent

arXiv:2604.13988v1 Announce Type: new Abstract: Objective: Investigate whether hypnogram 'realism' can be used to guide an unsupervised method for handling arbitrary types of signal degradation in mob

Using reasoning LLMs to extract SDOH events from clinical notes

ResearchDGX agent

arXiv:2604.13502v1 Announce Type: new Abstract: Social Determinants of Health (SDOH) refer to environmental, behavioral, and social conditions that influence how individuals live, work, and age. SDOH

Utilizing Inpainting for Keypoint Detection for Vision-Based Control of Robotic Manipulators

ResearchDGX agent

arXiv:2604.13309v1 Announce Type: new Abstract: In this paper we present a novel visual servoing framework to control a robotic manipulator in the configuration space by using purely natural visual fe

V2V With Audio File Lipsync?

Local AiDGX agent

This r/StableDiffusion post likely discusses how to perform video-to-video (V2V) generation with audio-driven lip synchronization, a workflow where an existing video is transformed so that a subject's

ValueGround: Evaluating Culture-Conditioned Visual Value Grounding in MLLMs

Model ReleasesDGX agent

arXiv:2604.06484v2 Announce Type: replace Abstract: Cultural values are expressed not only through language but also through visual scenes and everyday social practices. Yet existing evaluations of cu

Vectorizing Projection in Manifold-Constrained Motion Planning for Real-Time Whole-Body Control

ApplicationsDGX agent

arXiv:2604.13323v1 Announce Type: new Abstract: Many robot planning tasks require satisfaction of one or more constraints throughout the entire trajectory. For geometric constraints, manifold-constrai

Vercel Workflows is GA. Your code is the orchestrator. Ship agents, backends, or any long-running process without managing queues, retries, …

ToolsDGX agent

Vercel Workflows is GA. Your code is the orchestrator. Ship agents, backends, or any long-running process without managing queues, retries, or workers. https://vercel.com/blog/a-new-programming-model-

very excited for this one! a year ago, most of what was being traced to LangSmith was 'LLM apps'. now everything is becoming an agent. with …

AgentsDGX agent

very excited for this one! a year ago, most of what was being traced to LangSmith was 'LLM apps'. now everything is becoming an agent. with that shift, it's getting harder to know what your software i

VGGT-Segmentor: Geometry-Enhanced Cross-View Segmentation

Model ReleasesDGX agent

arXiv:2604.13596v1 Announce Type: new Abstract: Instance-level object segmentation across disparate egocentric and exocentric views is a fundamental challenge in visual understanding, critical for app

‘Vibe coding is fun, but is it safe?’: Oracle takes on the trust crisis at the core of AI development

ApplicationsDGX agent

Generative AI is democratizing application development, putting code-generation tools in the hands of virtually any developer. But as AI-assisted development accelerates, the enterprise technology ind

VibeFlow: Versatile Video Chroma-Lux Editing through Self-Supervised Learning

ResearchDGX agent

arXiv:2604.13425v1 Announce Type: new Abstract: Video chroma-lux editing, which aims to modify illumination and color while preserving structural and temporal fidelity, remains a significant challenge

ViBES: A Conversational Agent with Behaviorally-Intelligent 3D Virtual Body

Model ReleasesDGX agent

arXiv:2512.14234v2 Announce Type: replace Abstract: Human communication is inherently multimodal and social: words, prosody, and body language jointly carry intent. Yet most prior systems model human

VIGILant: an automatic classification pipeline for glitches in the Virgo detector

ResearchDGX agent

arXiv:2604.13687v1 Announce Type: cross Abstract: Glitches frequently contaminate data in gravitational-wave detectors, complicating the observation and analysis of astrophysical signals. This work in

Viktor Orbán’s electoral loss in Hungary is as much a defeat for Trump and JD Vance. 'Seldom have American leaders intervened so overtly in …

ResearchDGX agent

Viktor Orbán’s electoral loss in Hungary is as much a defeat for Trump and JD Vance. 'Seldom have American leaders intervened so overtly in a foreign election, and seldom has their preferred candidate

Vision-and-Language Navigation for UAVs: Progress, Challenges, and a Research Roadmap

AgentsDGX agent

arXiv:2604.13654v1 Announce Type: new Abstract: Vision-and-Language Navigation for Unmanned Aerial Vehicles (UAV-VLN) represents a pivotal challenge in embodied artificial intelligence, focused on ena

Visual Self-Fulfilling Alignment: Shaping Safety-Oriented Personas via Threat-Related Images

SafetyDGX agent

arXiv:2603.08486v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) face safety misalignment, where visual inputs enable harmful outputs. To address this, existing methods req

Visual Sparse Steering (VS2): Unsupervised Adaptation for Image Classification using Sparsity-Guided Steering Vectors

ResearchDGX agent

arXiv:2506.01247v2 Announce Type: replace Abstract: Steering vision foundation models at test time, without updating foundation-model weights or using labeled target data, is a desirable yet challengi

VLM Performance:Qwen3.6 is natively multimodal, and Qwen3.6-35B-A3B showcases perception and multimodal reasoning capabilities that far exce…

Model ReleasesDGX agent

VLM Performance:Qwen3.6 is natively multimodal, and Qwen3.6-35B-A3B showcases perception and multimodal reasoning capabilities that far exceed what its size would suggest, with only around 3 billion a

VLMs Need Words: Vision Language Models Ignore Visual Detail In Favor of Semantic Anchors

ResearchDGX agent

arXiv:2604.02486v2 Announce Type: replace-cross Abstract: Vision-language models (VLMs) have achieved impressive performance across a wide range of multimodal tasks. However, they often fail on tasks

Voice actors worldwide are mobilizing to protect their livelihoods and personality rights as Hollywood studios push AI dubbing to replace human performances (Rina Chandran/Rest of World)

IndustryDGX agent

Rina Chandran / Rest of World: Voice actors worldwide are mobilizing to protect their livelihoods and personality rights as Hollywood studios push AI dubbing to replace human performances — Voice AI t

VRAG-DFD: Verifiable Retrieval-Augmentation for MLLM-based Deepfake Detection

SafetyDGX agent

arXiv:2604.13660v1 Announce Type: new Abstract: In Deepfake Detection (DFD) tasks, researchers proposed two types of MLLM-based methods: complementary combination with small DFD detectors, or static f

WAI-ANIMA 1.0 released

Model ReleasesDGX agent

WAI-ANIMA 1.0 is a newly released Stable Diffusion checkpoint model from the WAI model family, likely combining elements of the WAI-Illustrious anime generation lineage with the Anima diffusion archit

WAIT WHAT?! 2-bit Qwen3.6-35B-A3B is lightning fast and it only needs 13 GB RAM. “did a complete repo bug hunt with evidence, repro, fixes, …

IndustryDGX agent

A developer successfully optimized Qwen 3.6-35B model to 2-bit quantization, achieving significant performance improvements with only 13GB RAM requirements while maintaining functionality. The work in

We are super excited to launch the in-app browser inside Codex with comment mode! View any web pages & iterate with your agent quickly with …

AgentsDGX agent

We are super excited to launch the in-app browser inside Codex with comment mode! View any web pages & iterate with your agent quickly with just point and click. Codex will automatically capture a scr

We comprehensively benchmarked Opus 4.7 on document understanding. We evaluated it through ParseBench - our comprehensive OCR benchmark for …

Model ReleasesDGX agent

We comprehensively benchmarked Opus 4.7 on document understanding. We evaluated it through ParseBench - our comprehensive OCR benchmark for enterprise documents where we evaluate tables, text, charts,

We fixed a bug where rate limits on Claude subscriptions weren't properly adjusted for long context requests in Opus 4.7. We've reset 5-hour…

Model ReleasesDGX agent

Anthropic fixed a bug in Claude Opus 4.7 where rate limits for paid subscriptions weren't correctly adjusted for requests using the model's extended context window capabilities. The fix involved reset

We partnered with University of Chicago economist @SuproteemSarkar to study how more capable models have changed the way people use Cursor. …

ToolsDGX agent

We partnered with University of Chicago economist @SuproteemSarkar to study how more capable models have changed the way people use Cursor. Across 500 teams, we find that developers are tackling more

We replicated Mythos findings in opencode using public models, not Anthropic's private stack. The moat is moving from model access to valida…

Model ReleasesDGX agent

We replicated Mythos findings in opencode using public models, not Anthropic's private stack. The moat is moving from model access to validation: finding vulnerability signal is getting cheaper; turni

Weakly-supervised Learning for Physics-informed Neural Motion Planning via Sparse Roadmap

Local AiDGX agent

arXiv:2604.13204v1 Announce Type: new Abstract: The motion planning problem requires finding a collision-free path between start and goal configurations in high-dimensional, cluttered spaces. Recent l

WebXSkill: Skill Learning for Autonomous Web Agents

AgentsDGX agent

arXiv:2604.13318v1 Announce Type: cross Abstract: Autonomous web agents powered by large language models (LLMs) have shown promise in completing complex browser tasks, yet they still struggle with lon

We’re back 🔥. Thrilled to be named once again to the @Forbes AI 50. The AI Native Cloud, built for the full AI lifecycle: fast inference, o…

ToolsDGX agent

We’re back 🔥. Thrilled to be named once again to the @Forbes AI 50. The AI Native Cloud, built for the full AI lifecycle: fast inference, open models, fine-tuning at scale. https://www.forbes.com/list

We're committed to bringing the power of vibe coding to the world. Join Brandon — Replit Fellow, AWS professional, and Stanford instructor —…

ToolsDGX agent

We're committed to bringing the power of vibe coding to the world. Join Brandon — Replit Fellow, AWS professional, and Stanford instructor — as he travels from city to city to connect with the builder

← Previous
1…12931294129512961297…1409
Next →