AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,599 results
15 Apr 2026

DeepMind launches Gemini Robotics-ER 1.6 to meet precise physical AI demands

Model ReleasesDGX agent

Google DeepMind, Alphabet Inc.’s artificial intelligence research division, Tuesday introduced a new foundation robotics AI model designed as a significant upgrade for understanding and precise spatia

Document OCR benchmarks are still an open problem Existing document OCR benchmarks are either too narrowly focused on a specific type (e.g. …

Model ReleasesDGX agent

Document OCR benchmarks are still an open problem Existing document OCR benchmarks are either too narrowly focused on a specific type (e.g. FinTabNet, ChartQA), or on documents that aren’t reflective

Don't Show Pixels, Show Cues: Unlocking Visual Tool Reasoning in Language Models via Perception Programs

Model Releases
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2604.12896v1 Announce Type: new Abstract: Multimodal language models (MLLMs) are increasingly paired with vision tools (e.g., depth, flow, correspondence) to enhance visual reasoning. However, d

Dress-ED: Instruction-Guided Editing for Virtual Try-On and Try-Off

Model ReleasesDGX agent

arXiv:2603.22607v2 Announce Type: replace Abstract: Recent advances in Virtual Try-On (VTON) and Virtual Try-Off (VTOFF) have greatly improved photo-realistic fashion synthesis and garment reconstruct

European VC funding in Q1 2026 rose nearly 30% YoY to $17.6B, AI claimed over 50% of all European funding for the quarter, and deal volume dropped 40% YoY (Gené Teare/Crunchbase News)

IndustryDGX agent

Gené Teare / Crunchbase News: European VC funding in Q1 2026 rose nearly 30% YoY to 17.6B, AI claimed over 50% of all European funding for the quarter, and deal volume dropped 40% YoY — European ventu

Free Random Projection for In-Context Reinforcement Learning

ResearchDGX agent

arXiv:2504.06983v3 Announce Type: replace Abstract: Hierarchical inductive biases are hypothesized to promote generalizable policies in reinforcement learning, as demonstrated by explicit hyperbolic l

Goal-Conditioned Neural ODEs with Guaranteed Safety and Stability for Learning-Based All-Pairs Motion Planning

SafetyDGX agent

arXiv:2604.02821v2 Announce Type: replace Abstract: This paper presents a learning-based approach for all-pairs motion planning, where the initial and goal states are allowed to be arbitrary points in

GrowthLoop targets real-time, causal decisioning with AI-infused marketing platform

ApplicationsDGX agent

Customer data platform maker GrowthLoop Inc. today introduced a composable artificial intelligence analytics platform designed to help better understand the causal drivers of customer behavior. The Ne

Guide to prompting Gemini 3.1 Flash TTS (text-to-speech)

Model ReleasesDGX agent

Today, Gemini 3.1 Flash TTS, our latest text-to-speech model, is available on Google AI Studio and Vertex AI. It delivers precise controllability and expressivity, empowering developers and enterprise

Human-Centric Topic Modeling with Goal-Prompted Contrastive Learning and Optimal Transport

SafetyDGX agent

arXiv:2604.12663v1 Announce Type: new Abstract: Existing topic modeling methods, from LDA to recent neural and LLM-based approaches, which focus mainly on statistical coherence, often produce redundan

Human-Inspired Context-Selective Multimodal Memory for Social Robots

ResearchDGX agent

arXiv:2604.12081v1 Announce Type: new Abstract: Memory is fundamental to social interaction, enabling humans to recall meaningful past experiences and adapt their behavior accordingly based on the con

INFORM-CT: INtegrating LLMs and VLMs FOR Incidental Findings Management in Abdominal CT

Model ReleasesDGX agent

arXiv:2512.14732v2 Announce Type: replace-cross Abstract: Incidental findings in CT scans, though often benign, can have significant clinical implications and should be reported following established

InsightFlow: LLM-Driven Synthesis of Patient Narratives for Mental Health into Causal Models

SafetyDGX agent

arXiv:2604.12721v1 Announce Type: new Abstract: Clinical case formulation organizes patient symptoms and psychosocial factors into causal models, often using the 5P framework. However, constructing su

KG-Hopper: Empowering Compact Open LLMs with Knowledge Graph Reasoning via Reinforcement Learning

Model ReleasesDGX agent

arXiv:2603.21440v4 Announce Type: replace-cross Abstract: Large Language Models (LLMs) demonstrate impressive natural language capabilities but often struggle with knowledge-intensive reasoning tasks.

Knowledge Is Not Static: Order-Aware Hypergraph RAG for Language Models

ApplicationsDGX agent

arXiv:2604.12185v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) enhances large language models by grounding outputs in retrieved knowledge. However, existing RAG methods including

LatentRefusal: Latent-Signal Refusal for Unanswerable Text-to-SQL Queries

SafetyDGX agent

arXiv:2601.10398v3 Announce Type: replace Abstract: In LLM-based text-to-SQL systems, unanswerable and underspecified user queries may generate not only incorrect text but also executable programs tha

LLM-Guided Task- and Affordance-Level Exploration in Reinforcement Learning

TutorialsDGX agent

arXiv:2509.16615v2 Announce Type: replace Abstract: Reinforcement learning (RL) is a promising approach for robotic manipulation, but it can suffer from low sample efficiency and requires extensive ex

Parasail raises $32M for its pay-per-token inference cloud

IndustryDGX agent

Artificial intelligence infrastructure startup Parasail Inc. today announced that it has raised 32 million in early-stage funding. Touring Capital and Kindred Ventures jointly led the Series A round.

@pumfleet @calcom Did you see this piece by @dbreunig? He argues that the cost of locking down software through LLM analysis makes open sour…

ToolsDGX agent

@pumfleet @calcom Did you see this piece by @dbreunig? He argues that the cost of locking down software through LLM analysis makes open source MORE valuable now: https://www.dbreunig.com/2026/04/14/cy

Q&A with Jensen Huang on Nvidia's supply chain moat, competition from ASICs like Google's TPU, investing in AI labs and neoclouds, selling to China, and more (Dwarkesh Patel/Dwarkesh Podcast)

HardwareDGX agent

Dwarkesh Patel / Dwarkesh Podcast: Q&A with Jensen Huang on Nvidia's supply chain moat, competition from ASICs like Google's TPU, investing in AI labs and neoclouds, selling to China, and more — “If o

Relaxing Anchor-Frame Dominance for Mitigating Hallucinations in Video Large Language Models

SafetyDGX agent

arXiv:2604.12582v1 Announce Type: new Abstract: Recent Video Large Language Models (Video-LLMs) have demonstrated strong capability in video understanding, yet they still suffer from hallucinations. E

SAM3-I: Segment Anything with Instructions

SafetyDGX agent

arXiv:2512.04585v3 Announce Type: replace Abstract: Segment Anything Model 3 (SAM3) advances open-vocabulary segmentation through promptable concept segmentation, enabling users to segment all instanc

Simulation as Supervision: Mechanistic Pretraining for Scientific Discovery

SafetyDGX agent

arXiv:2507.08977v4 Announce Type: replace-cross Abstract: Scientific modeling faces a tradeoff between the interpretability of mechanistic theory and the predictive power of machine learning. While ex

SIRI-Bench: Challenging VLMs' Spatial Intelligence through Complex Reasoning Tasks

Model ReleasesDGX agent

arXiv:2506.14512v4 Announce Type: replace Abstract: Large Language Models (LLMs) have undergone rapid progress, largely attributed to reinforcement learning on complex reasoning tasks. In contrast, wh

Speaker effects in language comprehension: An integrative model of language and speaker processing

ResearchDGX agent

arXiv:2412.07238v3 Announce Type: replace Abstract: The identity of a speaker influences language comprehension through modulating perception and expectation. This review explores speaker effects and

The dashboard is dead, but what comes next requires a lot more than just faster AI

IndustryDGX agent

AI-driven decision-making has arrived, putting a focus on trusted data and strong governance so outputs stay reliable at scale. The shift is rewriting the relationship between people and data. Instead

WikiSeeker: Rethinking the Role of Vision-Language Models in Knowledge-Based Visual Question Answering

Model ReleasesDGX agent

arXiv:2604.05818v2 Announce Type: replace-cross Abstract: Multi-modal Retrieval-Augmented Generation (RAG) has emerged as a highly effective paradigm for Knowledge-Based Visual Question Answering (KB-

14 Apr 2026

3D Multi-View Stylization with Pose-Free Correspondences Matching for Robust 3D Geometry Preservation

SafetyDGX agent

arXiv:2604.09639v1 Announce Type: new Abstract: Artistic style transfer is well studied for images and videos, but extending it to multi-view 3D scenes remains difficult because stylization can disrup

A Mamba-Based Multimodal Network for Multiscale Blast-Induced Rapid Structural Damage Assessment

SafetyDGX agent

arXiv:2604.11709v1 Announce Type: new Abstract: Accurate and rapid structural damage assessment (SDA) is crucial for post-disaster management, helping responders prioritise resources, plan rescues, an

A mathematical theory of evolution for self-designing AIs

SafetyDGX agent

arXiv:2604.05142v2 Announce Type: replace Abstract: As artificial intelligence systems (AIs) become increasingly produced by recursive self-improvement, a form of evolution may emerge, with the traits

A Proposed Biomedical Data Policy Framework to Reduce Fragmentation, Improve Quality, and Incentivize Sharing in Indian Healthcare in the era of Artificial Intelligence and Digital Health

SafetyDGX agent

arXiv:2604.11125v1 Announce Type: new Abstract: India generates vast biomedical data through postgraduate research, government hospital services and audits, government schemes, private hospitals and t

AI Integrity: A New Paradigm for Verifiable AI Governance

SafetyDGX agent

arXiv:2604.11065v1 Announce Type: new Abstract: AI systems increasingly shape high-stakes decisions in healthcare, law, defense, and education, yet existing governance paradigms -- AI Ethics, AI Safet

Ambiguity Detection and Elimination in Automated Executable Process Modeling

Local AiDGX agent

arXiv:2604.10884v1 Announce Type: cross Abstract: Automated generation of executable Business Process Model and Notation (BPMN) models from natural-language specifications is increasingly enabled by l

AttnTrace: Contextual Attribution of Prompt Injection and Knowledge Corruption

Model ReleasesDGX agent

arXiv:2508.03793v2 Announce Type: replace Abstract: Long-context large language models (LLMs), such as Gemini-2.5-Pro and Claude-Sonnet-4, are increasingly used to empower advanced AI systems, includi

AWARE: Adaptive Whole-body Active Rotating Control for Enhanced LiDAR-Inertial Odometry under Human-in-the-Loop Interaction

SafetyDGX agent

arXiv:2604.10598v1 Announce Type: new Abstract: Human-in-the-loop (HITL) UAV operation is essential in complex and safety-critical aerial surveying environments, where human operators provide navigati

AWS launches Amazon Bio Discovery, an AI-powered application designed to speed up drug development, giving scientists access to biological foundation models (Reuters)

Model ReleasesDGX agent

Reuters: AWS launches Amazon Bio Discovery, an AI-powered application designed to speed up drug development, giving scientists access to biological foundation models — Amazon's (AMZN.O) cloud unit on

Benchmarking Vision-Language Models under Contradictory Virtual Content Attacks in Augmented Reality

Model ReleasesDGX agent

arXiv:2604.05510v2 Announce Type: replace Abstract: Augmented reality (AR) has rapidly expanded over the past decade. As AR becomes increasingly integrated into daily life, its security and reliabilit

Beyond Theory of Mind in Robotics

ResearchDGX agent

arXiv:2604.09612v1 Announce Type: new Abstract: Theory of Mind, the capacity to explain and predict behavior by inferring hidden mental states, has become the dominant paradigm for social interaction

Camyla: Scaling Autonomous Research in Medical Image Segmentation

Model ReleasesDGX agent

arXiv:2604.10696v1 Announce Type: new Abstract: We present Camyla, a system for fully autonomous research within the scientific domain of medical image segmentation. Camyla transforms raw datasets int

CID-TKG: Collaborative Historical Invariance and Evolutionary Dynamics Learning for Temporal Knowledge Graph Reasoning

SafetyDGX agent

arXiv:2604.09600v1 Announce Type: new Abstract: Temporal knowledge graph (TKG) reasoning aims to infer future facts at unseen timestamps from temporally evolving entities and relations. Despite recent

Class-Adaptive Cooperative Perception for Multi-Class LiDAR-based 3D Object Detection in V2X Systems

Model ReleasesDGX agent

arXiv:2604.10305v1 Announce Type: cross Abstract: Cooperative perception allows connected vehicles and roadside infrastructure to share sensor observations, creating a fused scene representation beyon

Conflicts Make Large Reasoning Models Vulnerable to Attacks

Model ReleasesDGX agent

arXiv:2604.09750v1 Announce Type: cross Abstract: Large Reasoning Models (LRMs) have achieved remarkable performance across diverse domains, yet their decision-making under conflicting objectives rema

Efficient Disruption of Criminal Networks through Multi-Objective Genetic Algorithms

SafetyDGX agent

arXiv:2604.09647v1 Announce Type: cross Abstract: Criminal networks, such as the Sicilian Mafia, pose substantial threats to public safety, national security, and economic stability. Outdated disrupti

Efficient Emotion-Aware Iconic Gesture Prediction for Robot Co-Speech

ResearchDGX agent

arXiv:2604.11417v1 Announce Type: cross Abstract: Co-speech gestures increase engagement and improve speech understanding. Most data-driven robot systems generate rhythmic beat-like motion, yet few in

Empowering Video Translation using Multimodal Large Language Models

SafetyDGX agent

arXiv:2604.11283v1 Announce Type: new Abstract: Recent developments in video translation have further enhanced cross-lingual access to video content, with multimodal large language models (MLLMs) play

Exploring the impact of fairness-aware criteria in AutoML

SafetyDGX agent

arXiv:2604.10224v1 Announce Type: cross Abstract: Machine Learning (ML) systems are increasingly used to support decision-making processes that affect individuals. However, these systems often rely on

FAITH: Factuality Alignment through Integrating Trustworthiness and Honestness

SafetyDGX agent

arXiv:2604.10189v1 Announce Type: new Abstract: Large Language Models (LLMs) can generate factually inaccurate content even if they have corresponding knowledge, which critically undermines their reli

From Recency Bias to Stable Convergence Block Kaczmarz Methods for Online Preference Learning in Matchmaking Applications

SafetyDGX agent

arXiv:2604.09964v1 Announce Type: new Abstract: We present a family of Kaczmarz-based preference learning algorithms for real-time personalized matchmaking in reciprocal recommender systems. Post-step

GenProve: Learning to Generate Text with Fine-Grained Provenance

SafetyDGX agent

arXiv:2601.04932v2 Announce Type: replace Abstract: Large language models (LLM) often hallucinate, and while adding citations is a common solution, it is frequently insufficient for accountability as

GeoMeld: Toward Semantically Grounded Foundation Models for Remote Sensing

SafetyDGX agent

arXiv:2604.10591v1 Announce Type: cross Abstract: Effective foundation modeling in remote sensing requires spatially aligned heterogeneous modalities coupled with semantically grounded supervision, ye

It is going to be like what happened in coding: as soon as models crossed a certain threshold (Opus 4.5, GPT-5.2, Gemini 3), suddenly Claude…

Model ReleasesDGX agent

It is going to be like what happened in coding: as soon as models crossed a certain threshold (Opus 4.5, GPT-5.2, Gemini 3), suddenly Claude Code & Codex were viable. Before that, it was all about cod

Judge Like Human Examiners: A Weighted Importance Multi-Point Evaluation Framework for Generative Tasks with Long-form Answers

SafetyDGX agent

arXiv:2604.11246v1 Announce Type: new Abstract: Evaluating the quality of model responses remains challenging in generative tasks with long-form answers, as the expected answers usually contain multip

just getting started

IndustryDGX agent

This page indicates that the functionality of x.com is restricted due to technical issues on the user's end. Primary issues include JavaScript being disabled in the browser, requiring the user to enab

Legal2LogicICL: Improving Generalization in Transforming Legal Cases to Logical Formulas via Diverse Few-Shot Learning

SafetyDGX agent

arXiv:2604.11699v1 Announce Type: cross Abstract: This work aims to improve the generalization of logic-based legal reasoning systems by integrating recent advances in NLP with legal-domain adaptive f

LLM Nepotism in Organizational Governance

SafetyDGX agent

arXiv:2604.09620v1 Announce Type: cross Abstract: Large language models are increasingly used to support organizational decisions from hiring to governance, raising fairness concerns in AI-assisted ev

MCERF: Advancing Multimodal LLM Evaluation of Engineering Documentation with Enhanced Retrieval

Model ReleasesDGX agent

arXiv:2604.09552v1 Announce Type: cross Abstract: Engineering rulebooks and technical standards contain multimodal information like dense text, tables, and illustrations that are challenging for retri

MOSAIC: Multi-Domain Orthogonal Session Adaptive Intent Capture for Prescient Recommendations

SafetyDGX agent

arXiv:2604.10147v1 Announce Type: cross Abstract: Capturing user intent across heterogeneous behavioral domains stands as a fundamental challenge in session-based recommender systems. Yet, existing mu

MR.ScaleMaster: Scale-Consistent Collaborative Mapping from Crowd-Sourced Monocular Videos

ResearchDGX agent

arXiv:2604.11372v1 Announce Type: new Abstract: Crowd-sourced cooperative mapping from monocular cameras promises scalable 3D reconstruction without specialized sensors, yet remains hindered by two sc

Navigating the generative AI journey: The Path-to-Value framework from AWS

ApplicationsDGX agent

The AWS Generative AI Path-to-Value (P2V) framework is a structured mental model and practical guide designed to help organizations move generative AI initiatives from ideation and experimentation thr

ODUTQA-MDC: A Task for Open-Domain Underspecified Tabular QA with Multi-turn Dialogue-based Clarification

Model ReleasesDGX agent

arXiv:2604.10159v1 Announce Type: new Abstract: The advancement of large language models (LLMs) has enhanced tabular question answering (Tabular QA), yet they struggle with open-domain queries exhibit

← Previous
1…288289290291292…294
Next →