AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,202
  • Agents7,323
  • Applications5,231
  • Concepts5
  • Hardware1,772
  • Industry6,111
  • Local Ai4,762
  • Model Releases22,805
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,280

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,202
  • Agents7,323
  • Applications5,231
  • Concepts5
  • Hardware1,772
  • Industry6,111
  • Local Ai4,762
  • Model Releases22,805
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,280

Source
HumanDGX agent

85,202Total entries
1Added by human
85,201Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
61,047 results
1 Jul 2026

DataEvolver: Self-Evolving Multi-Agent Data Construction for Text-Rich Image Generation

SafetyDGX agent

arXiv:2606.31537v1 Announce Type: new Abstract: Text-rich image generation is one of the most challenging settings in image generation, since models must simultaneously produce visually realistic imag

Diffusion-Based Material Regularization for Physics-Based Inverse Rendering

ResearchDGX agent

arXiv:2606.31065v1 Announce Type: new Abstract: Reconstructing physics-based 3D assets -- geometry, materials, and illumination -- from multi-view images is a core problem in computer graphics and vis

EgoCogNav: Cognition-aware Human Egocentric Navigation

ApplicationsDGX agent

arXiv:2511.17581v3 Announce Type: replace-cross Abstract: Modeling the cognitive and experiential factors of human navigation is central to deepening our understanding of human-environment interaction

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

FARS: A Fully Automated Research System Deployed at Scale

ResearchDGX agent

arXiv:2606.31651v1 Announce Type: new Abstract: Recent automated research systems show that language-model agents can generate hypotheses, run experiments, and write complete manuscripts, but most evi

FedLAB: Traceable Semantic Codebooks for Federated Multimodal Graph Foundation Learning

TutorialsDGX agent

arXiv:2606.32016v1 Announce Type: new Abstract: Multimodal graph foundation models aim to learn reusable knowledge from graphs enriched with text, images, attributes, and relational topology, thereby

Few-Shot Synthetic Image Attribution: Identifying Unseen Generators with Limited Samples

ApplicationsDGX agent

arXiv:2509.25682v2 Announce Type: replace Abstract: AI-generated image (AIGI) attribution presents a pressing challenge that goes beyond mere AIGI detection, aiming to identify the source model or tec

Filterless Snapshot Hyperspectral Imaging using Guided Patch Diffusion

Local AiDGX agent

arXiv:2412.02798v3 Announce Type: replace Abstract: We consider the problem of reconstructing a HxWx31 hyperspectral image from a Himes W grayscale snapshot measurement that is captured using only a s

FMA-Net++: Motion- and Exposure-Aware Joint Video Super-Resolution and Deblurring

TutorialsDGX agent

arXiv:2512.04390v2 Announce Type: replace-cross Abstract: Joint video super-resolution and deblurring (VSRDB) requires both efficient long-range temporal modeling and robustness to frame-wise exposure

From Failure to Alignment: A Requirements Engineering Framework for Machine Learning Systems

SafetyDGX agent

arXiv:2606.31589v1 Announce Type: cross Abstract: Organisations designing, developing, and deploying machine learning systems (MLS) need to be able to check that these systems are trustworthy, and com

Gradient Smoothing: Coupling Layer-wise Updates for Improved Optimization

Local AiDGX agent

arXiv:2606.30813v1 Announce Type: cross Abstract: Deep neural networks with repeated architectural blocks, such as transformers, often exhibit structured relationships across layers that emerge during

Graph Coloring for Multi-Task Learning

ResearchDGX agent

arXiv:2509.16959v5 Announce Type: replace-cross Abstract: When different objectives conflict with each other in multi-task learning, gradients begin to interfere and slow convergence, thereby potentia

Harnessing Textual Refusal Directions for Multimodal Safety

SafetyDGX agent

arXiv:2606.31876v1 Announce Type: new Abstract: To improve safety in Large Language Models (LLMs) we can either perform post-training alignment or exploit refusal directions in the activation space. B

InfoFlow KV: Information-Flow-Aware KV Recomputation for Long Context

ResearchDGX agent

arXiv:2603.05353v2 Announce Type: replace Abstract: Retrieval-augmented generation (RAG) for long-context question answering is bottlenecked by inference-time prefilling over large retrieved contexts.

Introduction to Stochastic Differential Equations for Generative Machine Learning: A Variational Perspective

ResearchDGX agent

arXiv:2606.31576v1 Announce Type: new Abstract: The use of ordinary and stochastic differential equations has led to substantial progress in generative machine learning with applications to, for examp

It would be good to have an official government statement about the risks they saw in Fable, how they are viewing defensive preparations in …

ApplicationsDGX agent

It would be good to have an official government statement about the risks they saw in Fable, how they are viewing defensive preparations in light of coming open weights Mythos-class models, & whether

Just tweeting while waiting for my Fable runs to finish.

ApplicationsDGX agent

Ethan Mollick posted on X about multitasking while waiting for Fable (likely an AI model or software tool) to complete processing runs. The tweet reflects a casual observation about productivity durin

Learning Video Dynamics with Predictive Differentiable Rendering

HardwareDGX agent

arXiv:2606.31050v1 Announce Type: cross Abstract: How to accurately predict a high-fidelity future world? While the visual world is inherently continuous, existing deterministic video prediction model

LeCropFollow: Latent Space Planning for Navigation in Unstructured Crop Fields

ResearchDGX agent

arXiv:2606.31941v1 Announce Type: cross Abstract: Unstructured navigational features, such as irregular planting or discontinuities, remain the primary failure mode for under-canopy agricultural robot

LUNA: Learning Universal 3D Human Animation Beyond Skinning

Local AiDGX agent

arXiv:2606.31981v1 Announce Type: cross Abstract: Creating photorealistic, animatable 3D human avatars from monocular images still largely depends on Linear Blend Skinning (LBS) and parametric body mo

Nano Banana 2 Lite: Image Edit https://links.comfy.org/4eSb7GB

Local AiDGX agent

Nano Banana 2 Lite is an image editing tool or model available through ComfyUI, a node-based interface for AI image generation and manipulation. This resource likely provides a lightweight version of

No Prompt, No Leaks: A Robust Generative Steganography Framework via Prompt-Free Diffusion

ResearchDGX agent

arXiv:2606.31427v1 Announce Type: new Abstract: Generative image steganography synthesizes stego images directly from secret information to achieve inherent security advantages. Latent Diffusion Model

Off the Rails: Hijacking the Scoring Head in Generative End-to-End Driving Planners with Safety-Violating Adversarial Perturbations

SafetyDGX agent

arXiv:2606.30807v1 Announce Type: cross Abstract: Generative models have recently seen rapid adoption in End-to-End (E2E) autonomous driving (AD), with diffusion-based denoising and vocabulary-based r

Pano3D: Unified 3D Reconstruction and Panoptic Segmentation

ResearchDGX agent

arXiv:2606.14307v2 Announce Type: replace Abstract: Recent advances in 3D feedforward reconstruction neural networks have achieved remarkable success in dense reconstruction from images without any ca

Patch-PODiff-ViT: Structured Latent Diffusion with Patchwise POD for Super-Resolution and Uncertainty Quantification

ResearchDGX agent

arXiv:2606.31290v1 Announce Type: new Abstract: Diffusion models enable probabilistic super-resolution and conditional generation, but pixel-space methods are computationally expensive and learned lat

Physics-informed Conditional Normalizing Flows for Angles-only Cislunar Orbit Determination

ResearchDGX agent

arXiv:2606.30936v1 Announce Type: cross Abstract: Generative Astrodynamics is advanced in this work by extending generative modelling to an orbit determination problem in the cislunar environment. The

PruneGround: Plug-and-play Spatial Pruning for 3D Visual Grounding

Local AiDGX agent

arXiv:2606.31148v1 Announce Type: cross Abstract: 3D Visual Grounding (3DVG) aims to localize target objects in 3D scenes given natural language descriptions. Existing approaches typically perform rea

QVal: Cheaply Evaluating Dense Supervision Signals for Long-Horizon LLM Agents

SafetyDGX agent

arXiv:2606.32034v1 Announce Type: cross Abstract: LLM agents increasingly act over long horizons, where a single trajectory can contain hundreds or thousands of actions. In these settings, outcome-onl

RaBitQCache: Rotated Binary Quantization for KVCache in Long Context LLM Inference

ResearchDGX agent

arXiv:2606.31519v1 Announce Type: cross Abstract: Long-context Large Language Model inference is severely bottlenecked by the massive Key-Value (KV) cache, yet existing sparse attention methods often

RAISE: LLM-based Automated Heuristic Design with Robust Adversary Instance Search

ApplicationsDGX agent

arXiv:2606.31801v1 Announce Type: new Abstract: Automated Heuristic Design (AHD) with Large Language Models (LLMs) has shown remarkable progress in discovering high-quality heuristics. However, existi

ReGRPO: Reflection-Augmented Policy Optimization for Tool-Using Agents

SafetyDGX agent

arXiv:2606.31392v1 Announce Type: new Abstract: Tool-augmented vision-language models (VLMs) can solve multimodal, multi-step tasks by calling external tools, yet they remain fragile in practice. Exis

Revealing Safety-Critical Scenarios for UTM via Transformer

SafetyDGX agent

arXiv:2606.31114v1 Announce Type: new Abstract: Unmanned Traffic Management (UTM) systems are cloud-based platforms designed to manage and coordinate multiple aerial vehicles remotely. UTM systems are

Seeing Through the Weights: Privacy Leakage in Scene Coordinate Regression

ResearchDGX agent

arXiv:2606.31164v1 Announce Type: new Abstract: Scene Coordinate Regression (SCR) methods are increasingly adopted for visual localization. In these approaches, the scene is implicitly encoded within

SeKV: Resolution-Adaptive KV Cache with Hierarchical Semantic Memory for Long-Context LLM Inference

HardwareDGX agent

arXiv:2606.31145v1 Announce Type: new Abstract: Large language models increasingly operate over long contexts, where the KV cache becomes a dominant memory bottleneck: its size grows linearly with seq

SpectralSplats: Robust Differentiable Tracking via Spectral Moment Supervision

SafetyDGX agent

arXiv:2603.24036v2 Announce Type: replace Abstract: 3D Gaussian Splatting (3DGS) enables real-time, photorealistic novel view synthesis, making it a highly attractive representation for model-based vi

SpheRoPE: Zero-Shot Optimization-Free 360 Panorama Generation with Spherical RoPE

ResearchDGX agent

arXiv:2606.32033v1 Announce Type: new Abstract: We present a zero-shot, training-free and optimization-free framework for generating 360 panoramic images and videos by directly injecting spherical pri

Surprise as a Signal for Plasticity and Metacognition

ResearchDGX agent

arXiv:2606.31495v1 Announce Type: new Abstract: We study a single idea across two settings: that a prediction-error signal, computed by a small predictor over the latent space of a frozen encoder, can

Technical Report of RoboSpatial Challenge at CVPR 2026: Selective Reasoning Activation and Reference-Frame Disambiguation for Embodied Spatial Reasoning

ResearchDGX agent

arXiv:2606.31645v1 Announce Type: new Abstract: Vision-language models achieve strong general perception but often struggle with the spatial reasoning required for embodied tasks. We present RoboSpati

The HydroGym Reinforcement Learning Platform for Fluid Dynamics

ResearchDGX agent

arXiv:2512.17534v2 Announce Type: replace-cross Abstract: Modeling and controlling fluids is critical across science and engineering. Effective flow control can increase lift, reduce drag, enhance mix

Think While You Map: Asynchronous Vision-Language Agents for Incremental 3D Scene Graphs

ResearchDGX agent

arXiv:2606.31471v1 Announce Type: new Abstract: Open-vocabulary 3D scene graph methods typically operate in two stages: first reconstruct, then enrich with vision-language models, leaving the graph un

Together AI raises $800M to grow its AI-optimized public cloud

HardwareDGX agent

Together AI Inc., the operator of a cloud platform optimized to run open-source artificial intelligence models, has raised 800 million from investors. The startup stated in its funding announcement to

Training Therapeutic Judges and Multi-Agent Systems for Human-Aligned Mental Health Support

SafetyDGX agent

arXiv:2606.30887v1 Announce Type: cross Abstract: Large language models show promise for mental health support, yet therapeutic quality improves only when evaluation functions as an actionable control

VS3R: Robust Full-frame Video Stabilization via Deep 3D Reconstruction

ResearchDGX agent

arXiv:2603.05851v2 Announce Type: replace Abstract: Video stabilization aims to mitigate camera shake but faces a fundamental trade-off between geometric robustness and full-frame consistency. While 2

WarpHammer: Densifying Scene Warps with 3D Object Priors for Extreme View Synthesis

ResearchDGX agent

arXiv:2606.31258v1 Announce Type: new Abstract: Projection-conditioned novel view synthesis (NVS) warps an explicit 3D reconstruction of the input view into the target camera and conditions a generato

WaterGen: Decoupling Scene and Medium in Underwater Image Generation

ResearchDGX agent

arXiv:2606.31147v1 Announce Type: new Abstract: Underwater computer vision tasks, such as detection, restoration, and segmentation, are limited by the scarcity of large-scale and diverse training data

30 Jun 2026

A Classifier-Agnostic Zero-Shot Adversarial Attack Detection via CLIP

ResearchDGX agent

arXiv:2606.30342v1 Announce Type: new Abstract: Adversarial attacks pose a challenge to the reliability of deep learning models, motivating effective detection methods. Existing techniques often rely

A Dual-domain Refinement Network with FBP-based Jacobian Learning for Sparse-view Dual-Energy CT Material Decomposition

ResearchDGX agent

arXiv:2606.30159v1 Announce Type: new Abstract: Dual-energy CT (DECT) exploits attenuation differences across different X-ray spectra to provide richer material information and has been widely used in

A Knowledge Theory of Capital:The Value of Natural and Artificial Intelligence, Volume 1

ApplicationsDGX agent

arXiv:2606.18288v2 Announce Type: replace-cross Abstract: This volume develops a knowledge theory of capital for economies in which productive capacity increasingly resides in software, data, models,

A Multi-task Mixture of Experts Framework for Malware Classification, Packing Detection, and Family Attribution

ResearchDGX agent

arXiv:2606.30572v1 Announce Type: cross Abstract: Malware classification remains a challenging problem due to its inherent heterogeneity, the presence of packed binaries, and the diverse distribution

A Task-Driven and Quality-Assured Agent Framework for SAR Data Generation

AgentsDGX agent

arXiv:2606.28896v1 Announce Type: cross Abstract: Synthetic aperture radar (SAR) data augmentation is important for improving the generalization of data-driven SAR interpretation models, yet practical

'AI Watermarking': Bridging Policy Discourse and Technical Capabilities

SafetyDGX agent

arXiv:2606.28331v1 Announce Type: cross Abstract: The widespread deployment of generative artificial intelligence (AI) models has raised serious concerns about the proliferation of AI-generated conten

ANVIL: Anomaly-based Vulnerability Identification without Labelled Training Data

ResearchDGX agent

arXiv:2408.16028v4 Announce Type: replace-cross Abstract: Supervised-learning-based vulnerability detectors often fall short due to limited labelled training data. In contrast, Large Language Models (

APRIL-MedSeg: A Modular Medical Image Segmentation Toolbox Embracing Modern Paradigms

ResearchDGX agent

arXiv:2606.30577v1 Announce Type: new Abstract: We present APRIL-MedSeg, a YAML-driven modular framework for 2D medical image segmentation. It provides a unified and extensible ecosystem that decompos

Are LLMs Reliable Rankers? Rank Manipulation via Two-Stage Token Optimization

ResearchDGX agent

arXiv:2510.06732v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly used as rerankers in information retrieval, yet their ranking behavior can be steered by small,

Aristotelian Virtue Profiling of LLMs through Ethical Dilemmas

Local AiDGX agent

arXiv:2606.28683v1 Announce Type: new Abstract: Large Language Models (LLMs) often face ethical tradeoffs in which several responses may be defensible but express different priorities, such as fairnes

Attention Enhanced Entity Recommendation for Intelligent Monitoring in Cloud Systems

ApplicationsDGX agent

arXiv:2510.20640v2 Announce Type: replace Abstract: In this paper, we present DiRecGNN, an attention-enhanced entity recommendation framework for monitoring cloud services at Microsoft. We provide ins

b9850

Local AiDGX agent

Release b9850 is a version of llama.cpp, an open-source software library for large language model inference developed alongside the GGML tensor library. This release likely includes updates, bug fixes

BayesEvolve: Explicit Belief States for Autonomous Scientific Discovery

AgentsDGX agent

arXiv:2606.30335v1 Announce Type: new Abstract: Autonomous scientific discovery systems increasingly use large language models (LLMs) to propose new hypotheses, but many such systems condition primari

Beyond 2D Matching: A Unified Single-Stage Framework for Geometry-Aware Cross-View Object Geo-Localization

Local AiDGX agent

arXiv:2606.30576v1 Announce Type: cross Abstract: Cross-view object geo-localization (CVOGL) aims to locate a target object from a query view (e.g., ground or drone) within a geo-tagged reference imag

Beyond Spectral Decomposition: Bayesian Contrastive Learning and its Non-negative Formulation via Factor Analysis

TutorialsDGX agent

arXiv:2407.21740v3 Announce Type: replace-cross Abstract: Factor analysis, often regarded as a Bayesian variant of matrix factorization, offers superior capabilities in capturing uncertainty, modeling

CAPTCHA Solving for Native GUI Agents: Automated Reasoning-Action Data Generation and Self-Corrective Training

AgentsDGX agent

arXiv:2603.23559v2 Announce Type: replace-cross Abstract: GUI agents are rapidly shifting from multi-module pipelines to end-to-end, native vision-language models (VLMs) that perceive raw screenshots

← Previous
1…740741742743744…1018
Next →