AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
83,773 results
13 Apr 2026

ELT: Elastic Looped Transformers for Visual Generation

Model ReleasesDGX agent

arXiv:2604.09168v1 Announce Type: new Abstract: We introduce Elastic Looped Transformers (ELT), a highly parameter-efficient class of visual generative models based on a recurrent transformer architec

EMA Is Not All You Need: Mapping the Boundary Between Structure and Content in Recurrent Context

Model ReleasesDGX agent

arXiv:2604.08556v1 Announce Type: cross Abstract: What exactly do efficient sequence models gain over simple temporal averaging? We use exponential moving average (EMA) traces, the simplest recurrent

Emad Mostaque built Stable Diffusion. I asked him what the trillion-dollar AI labs won't say publicly. @EMostaque's answer: 'Soon, adding a …

IndustryDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub

Emad Mostaque built Stable Diffusion. I asked him what the trillion-dollar AI labs won't say publicly. @EMostaque's answer: 'Soon, adding a human to the best AI team will make it worse. Not slower. Ne

EmoCtrl: Controllable Emotional Image Content Generation

SafetyDGX agent

arXiv:2512.22437v2 Announce Type: replace Abstract: An image conveys meaning through both its visual content and emotional tone, jointly shaping human perception. We introduce Controllable Emotional I

EngageTriBoost: Predictive Modeling of User Engagement in Digital Mental Health Intervention Using Explainable Machine Learning

ResearchDGX agent

arXiv:2604.08589v1 Announce Type: new Abstract: Mental health challenges among young adults, are on the rise, necessitating effective solutions such as digital mental health interventions (DMHIs). Des

Enhanced Self-Supervised Multi-Image Super-Resolution for Camera Array Images

ApplicationsDGX agent

arXiv:2604.06816v2 Announce Type: replace-cross Abstract: Conventional multi-image super-resolution (MISR) methods, such as burst and video SR, rely on sequential frames from a single camera. Conseque

Enhancing LLM Problem Solving via Tutor-Student Multi-Agent Interaction

Model ReleasesDGX agent

arXiv:2604.08931v1 Announce Type: new Abstract: Human cognitive development is shaped not only by individual effort but by structured social interaction, where role-based exchanges such as those betwe

Enhancing the Safety of Medical Vision-Language Models by Synthetic Demonstrations

SafetyDGX agent

arXiv:2506.09067v2 Announce Type: replace-cross Abstract: Generative medical vision-language models~(Med-VLMs) are primarily designed to generate complex textual information~(e.g., diagnostic reports)

Enterprises power agentic workflows in Cloudflare Agent Cloud with OpenAI

Model ReleasesDGX agent

Cloudflare and OpenAI have partnered to create the Cloudflare Agent Cloud, a platform that enables enterprises to build and deploy agentic AI workflows at scale using OpenAI's models and APIs. The int

Envisioning the Future, One Step at a Time

Model ReleasesDGX agent

arXiv:2604.09527v1 Announce Type: cross Abstract: Accurately anticipating how complex, diverse scenes will evolve requires models that represent uncertainty, simulate along extended interaction chains

EpiAgent: An Agent-Centric System for Ancient Inscription Restoration

AgentsDGX agent

arXiv:2604.09367v1 Announce Type: new Abstract: Ancient inscriptions, as repositories of cultural memory, have suffered from centuries of environmental and human-induced degradation. Restoring their i

EquiformerV3: Scaling Efficient, Expressive, and General SE(3)-Equivariant Graph Attention Transformers

ResearchDGX agent

arXiv:2604.09130v1 Announce Type: cross Abstract: As SE(3)-equivariant graph neural networks mature as a core tool for 3D atomistic modeling, improving their efficiency, expressivity, and physical con

@ErnstRoets All the politicians in South Africa who push their viciously racist laws should be sanctioned, barred from travel, declared crim…

IndustryDGX agent

Elon Musk made a post on X (formerly Twitter) directed at Ernst Roets, a South African civil rights activist, expressing strong opposition to South African politicians who support race-based legislati

@ESYudkowsky My horrendous nightmare of a political lifecycle, ladies and gentlemen and others.

SafetyDGX agent

Connor Leahy shared a post on X (formerly Twitter) quoting or referencing Eliezer Yudkowsky's account (@ESYudkowsky), describing what he characterizes as a 'horrendous nightmare of a political lifecyc

EthicMind: A Risk-Aware Framework for Ethical-Emotional Alignment in Multi-Turn Dialogue

SafetyDGX agent

arXiv:2604.09265v1 Announce Type: new Abstract: Intelligent dialogue systems are increasingly deployed in emotionally and ethically sensitive settings, where failures in either emotional attunement or

Event-Driven Temporal Graph Networks for Asynchronous Multi-Agent Cyber Defense in NetForge_RL

AgentsDGX agent

arXiv:2604.09523v1 Announce Type: new Abstract: The transition of Multi-Agent Reinforcement Learning (MARL) policies from simulated cyber wargames to operational Security Operations Centers (SOCs) is

Every Response Counts: Quantifying Uncertainty of LLM-based Multi-Agent Systems through Tensor Decomposition

AgentsDGX agent

arXiv:2604.08708v1 Announce Type: cross Abstract: While Large Language Model-based Multi-Agent Systems (MAS) consistently outperform single-agent systems on complex tasks, their intricate interactions

Everyone who think vision is solved should look at some enterprise documents

Model ReleasesDGX agent

Everyone who think vision is solved should look at some enterprise documents We’re open sourcing the first document OCR benchmark for the agentic era, ParseBench. Document parsing is the foundation of

Evidential Transformation Network: Turning Pretrained Models into Evidential Models for Post-hoc Uncertainty Estimation

ApplicationsDGX agent

arXiv:2604.08627v1 Announce Type: cross Abstract: Pretrained models have become standard in both vision and language, yet they typically do not provide reliable measures of confidence. Existing uncert

EVOKE: Emotion Vocabulary Of Korean and English

ResearchDGX agent

arXiv:2602.10414v2 Announce Type: replace Abstract: This paper introduces EVOKE (Emotion Vocabulary of Korean and English), a Korean-English parallel dataset of emotion words. The dataset offers compr

EvoLen: Evolution-Guided Tokenization for DNA Language Model

SafetyDGX agent

arXiv:2604.08698v1 Announce Type: new Abstract: Tokens serve as the basic units of representation in DNA language models (DNALMs), yet their design remains underexplored. Unlike natural language, DNA

Evolutionary Optimization Trumps Adam Optimization on Embedding Space Exploration

SafetyDGX agent

arXiv:2511.03913v2 Announce Type: replace-cross Abstract: Deep diffusion models have revolutionized image generation by producing high-quality outputs. However, achieving specific objectives with thes

EVs ranked by total unit sold in the U.S. in Q1 2026. 1) Tesla Model Y: 78,591 2) Tesla Model 3: 31,672 3) Toyota bZ: 10,029 4) Hyundai Ioni…

IndustryDGX agent

EVs ranked by total unit sold in the U.S. in Q1 2026. 1) Tesla Model Y: 78,591 2) Tesla Model 3: 31,672 3) Toyota bZ: 10,029 4) Hyundai Ioniq 5: 9,790 5) Chevrolet Equinox EV: 9,589 6) Rivian R1S: 5,4

EXAONE 4.5 Technical Report

Model ReleasesDGX agent

arXiv:2604.08644v1 Announce Type: new Abstract: This technical report introduces EXAONE 4.5, the first open-weight vision language model released by LG AI Research. EXAONE 4.5 is architected by integr

Excited to have @davidgomes of @cursor_ai take the AIE Miami stage! In his talk 'IDEs are dead. Long live IDEs' he questions whether IDEs ar…

ToolsDGX agent

Excited to have @davidgomes of @cursor_ai take the AIE Miami stage! In his talk 'IDEs are dead. Long live IDEs' he questions whether IDEs are truly dying, arguing instead that they’re evolving. Don't

Exploiting Web Search Tools of AI Agents for Data Exfiltration

ResearchDGX agent

arXiv:2510.09093v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are now routinely used to autonomously execute complex tasks, from natural language processing to dynamic workflo

Explorable Theorems: Making Written Theorems Explorable by Grounding Them in Formal Representations

ResearchDGX agent

arXiv:2604.02598v2 Announce Type: replace-cross Abstract: LLM-generated explanations can make technical content more accessible, but there is a ceiling on what they can support interactively. Because

Exploring Cross-lingual Latent Transplantation: Mutual Opportunities and Open Challenges

ResearchDGX agent

arXiv:2412.12686v3 Announce Type: replace Abstract: Current large language models (LLMs) often exhibit imbalances in multilingual capabilities and cultural adaptability, largely attributed to their En

Exploring Teachers' Perspectives on Using Conversational AI Agents for Group Collaboration

SafetyDGX agent

arXiv:2602.07142v2 Announce Type: replace-cross Abstract: Collaboration is a cornerstone of 21st-century learning, yet teachers continue to face challenges in supporting productive peer interaction. E

Exploring the new `servo` crate

Model ReleasesDGX agent

Research: Exploring the new `servo` crate In Servo is now available on crates.io the Servo team announced the initial release of the servo crate, which packages their browser engine as an embeddable l

Extrapolating Volition with Recursive Information Markets

SafetyDGX agent

arXiv:2604.08606v1 Announce Type: cross Abstract: One of the impediments to the efficiency of information markets is the inherent information asymmetry present in them, exacerbated by the 'buyer's ins

Face vs body Zit Lora

Local AiDGX agent

This Reddit post from r/StableDiffusion discusses a community comparison or showcase involving a 'Zit LoRA' — a Stable Diffusion LoRA model — examining how it performs differently when applied to face

FaceLiVTv2: An Improved Hybrid Architecture for Efficient Mobile Face Recognition

Local AiDGX agent

arXiv:2604.09127v1 Announce Type: new Abstract: Lightweight face recognition is increasingly important for deployment on edge and mobile devices, where strict constraints on latency, memory, and energ

Facet-Level Tracing of Evidence Uncertainty and Hallucination in RAG

Model ReleasesDGX agent

arXiv:2604.09174v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) aims to reduce hallucination by grounding answers in retrieved evidence, yet hallucinated answers remain common eve

FashionStylist: An Expert Knowledge-enhanced Multimodal Dataset for Fashion Understanding

Model ReleasesDGX agent

arXiv:2604.09249v1 Announce Type: new Abstract: Fashion understanding requires both visual perception and expert-level reasoning about style, occasion, compatibility, and outfit rationale. However, ex

Fast-dVLM: Efficient Block-Diffusion VLM via Direct Conversion from Autoregressive VLM

AgentsDGX agent

arXiv:2604.06832v2 Announce Type: replace Abstract: Vision-language models (VLMs) predominantly rely on autoregressive decoding, which generates tokens one at a time and fundamentally limits inference

Fast Model-guided Instance-wise Adaptation Framework for Real-world Pansharpening with Fidelity Constraints

HardwareDGX agent

arXiv:2604.08903v1 Announce Type: new Abstract: Pansharpening aims to generate high-resolution multispectral (HRMS) images by fusing low-resolution multispectral (LRMS) and high-resolution panchromati

FDIF: Formula-Driven supervised Learning with Implicit Functions for 3D Medical Image Segmentation

ResearchDGX agent

arXiv:2603.23199v2 Announce Type: replace Abstract: Deep learning-based 3D medical image segmentation methods relies on large-scale labeled datasets, yet acquiring such data is difficult due to privac

Feature-Label Modal Alignment for Robust Partial Multi-Label Learning

SafetyDGX agent

arXiv:2604.09064v1 Announce Type: new Abstract: In partial multi-label learning (PML), each instance is associated with a set of candidate labels containing both ground-truth and noisy labels. The pre

Few-Shot Contrastive Adaptation for Audio Abuse Detection in Low-Resource Indic Languages

ResearchDGX agent

arXiv:2604.09094v1 Announce Type: cross Abstract: Abusive speech detection is becoming increasingly important as social media shifts towards voice-based interaction, particularly in multilingual and l

Few-Shot Personalized Age Estimation

Model ReleasesDGX agent

arXiv:2604.09125v1 Announce Type: new Abstract: Existing age estimation methods treat each face as an independent sample, learning a global mapping from appearance to age. This ignores a well-document

Filing: Anthropic hired Ballard Partners, a lobbying firm with strong Trump administration ties, days after the DOD designated the startup a supply chain risk (Bloomberg)

IndustryDGX agent

Bloomberg: Filing: Anthropic hired Ballard Partners, a lobbying firm with strong Trump administration ties, days after the DOD designated the startup a supply chain risk — Anthropic PBC hired the lobb

Fine-Grained Action Segmentation for Renorrhaphy in Robot-Assisted Partial Nephrectomy

Model ReleasesDGX agent

arXiv:2604.09051v1 Announce Type: new Abstract: Fine-grained action segmentation during renorrhaphy in robot-assisted partial nephrectomy requires frame-level recognition of visually similar suturing

Fine-tuning is Not Enough: A Parallel Framework for Collaborative Imitation and Reinforcement Learning in End-to-end Autonomous Driving

Model ReleasesDGX agent

arXiv:2603.13842v3 Announce Type: replace-cross Abstract: End-to-end autonomous driving is typically built upon imitation learning (IL), yet its performance is constrained by the quality of human demo

Finite-Sample Analysis of Nonlinear Independent Component Analysis:Sample Complexity and Identifiability Bounds

Model ReleasesDGX agent

arXiv:2604.08850v1 Announce Type: new Abstract: Independent Component Analysis (ICA) is a fundamental unsupervised learning technique foruncovering latent structure in data by separating mixed signals

FIRE-CIR: Fine-grained Reasoning for Composed Fashion Image Retrieval

Model ReleasesDGX agent

arXiv:2604.09114v1 Announce Type: new Abstract: Composed image retrieval (CIR) aims to retrieve a target image that depicts a reference image modified by a textual description. While recent vision-lan

Fisher-Geometric Diffusion in Stochastic Gradient Descent: Optimal Rates, Oracle Complexity, and Information-Theoretic Limits

ResearchDGX agent

arXiv:2603.02417v3 Announce Type: replace-cross Abstract: Classical stochastic-approximation analyses treat the covariance of stochastic gradients as an exogenous modeling input. We show that under ex

FIT-GNN: Faster Inference Time for GNNs that 'FIT' in Memory Using Coarsening

Model ReleasesDGX agent

arXiv:2410.15001v5 Announce Type: replace Abstract: Scalability of Graph Neural Networks (GNNs) remains a significant challenge. To tackle this, methods like coarsening, condensation, and computation

FlashLips: 100-FPS Mask-Free Latent Lip-Sync using Reconstruction Instead of Diffusion or GANs

HardwareDGX agent

arXiv:2512.20033v2 Announce Type: replace Abstract: We present FlashLips, a two-stage, mask-free lip-sync system that decouples lips control from rendering and achieves real-time performance, with our

FluidFlow: a flow-matching generative model for fluid dynamics surrogates on unstructured meshes

Model ReleasesDGX agent

arXiv:2604.08586v1 Announce Type: cross Abstract: Computational fluid dynamics (CFD) provides high-fidelity simulations of fluid flows but remains computationally expensive for many-query applications

Folks, this is not Jevon's Paradox. this is just normal supply and demand. It turns out the utility of AI is high enough that people have hi…

ApplicationsDGX agent

Folks, this is not Jevon's Paradox. this is just normal supply and demand. It turns out the utility of AI is high enough that people have high demand, which is outstripping supply (so prices will go u

For the first time, 50% of employed US adults say they use AI at work a few times per year or more; leaders are more likely to see AI's impact as positive (Andy Kemp/Gallup)

IndustryDGX agent

Andy Kemp / Gallup: For the first time, 50% of employed US adults say they use AI at work a few times per year or more; leaders are more likely to see AI's impact as positive — Employees report produc

FP8-RL: A Practical and Stable Low-Precision Stack for LLM Reinforcement Learning

SafetyDGX agent

arXiv:2601.18150v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) for large language models (LLMs) is increasingly bottlenecked by rollout (generation), where long output sequence

Free AI Voice Cloning with Qwen3 TTS — Google Colab Notebook (works on free tier, no GPU needed)

HardwareDGX agent

This Reddit post shares a Google Colab notebook that enables free AI voice cloning using Alibaba's Qwen3-TTS, which offers voice cloning, voice design, and ultra-high-quality human-like speech generat

Frequency-Enhanced Diffusion Models: Curriculum-Guided Semantic Alignment for Zero-Shot Skeleton Action Recognition

SafetyDGX agent

arXiv:2604.09063v1 Announce Type: cross Abstract: Human action recognition is pivotal in computer vision, with applications ranging from surveillance to human-robot interaction. Despite the effectiven

From Business Events to Auditable Decisions: Ontology-Governed Graph Simulation for Enterprise AI

Model ReleasesDGX agent

arXiv:2604.08603v1 Announce Type: new Abstract: Existing LLM-based agent systems share a common architectural failure: they answer from the unrestricted knowledge space without first simulating how ac

From Dispersion to Attraction: Spectral Dynamics of Hallucination Across Whisper Model Scales

SafetyDGX agent

arXiv:2604.08591v1 Announce Type: cross Abstract: Hallucinations in large ASR models present a critical safety risk. In this work, we propose the extit{Spectral Sensitivity Theorem}, which predicts a

From Frames to Events: Rethinking Evaluation in Human-Centric Video Anomaly Detection

Local AiDGX agent

arXiv:2604.09327v1 Announce Type: new Abstract: Pose-based Video Anomaly Detection (VAD) has gained significant attention for its privacy-preserving nature and robustness to environmental variations.

from my experience, even the best models (Opus 4.6, 5.4 xhigh / 5.3 codex) cannot write good code today without an amount of work that is eq…

TutorialsDGX agent

from my experience, even the best models (Opus 4.6, 5.4 xhigh / 5.3 codex) cannot write good code today without an amount of work that is equivalent to just doing the work myself am excited for a worl

From Navigation to Refinement: Revealing the Two-Stage Nature of Flow-based Diffusion Models through Oracle Velocity

ResearchDGX agent

arXiv:2512.02826v3 Announce Type: replace-cross Abstract: Flow-based diffusion models have emerged as a leading paradigm for training generative models across images and videos. However, their memoriz

← Previous
1…13391340134113421343…1397
Next →