AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,617
  • Agents7,497
  • Applications5,365
  • Concepts5
  • Hardware1,816
  • Industry6,151
  • Local Ai4,900
  • Model Releases23,593
  • Research19,967
  • Safety13,267
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,617
  • Agents7,497
  • Applications5,365
  • Concepts5
  • Hardware1,816
  • Industry6,151
  • Local Ai4,900
  • Model Releases23,593
  • Research19,967
  • Safety13,267
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
HumanDGX agent

87,617Total entries
1Added by human
87,616Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,987 results
13 Apr 2026

TinyNeRV: Compact Neural Video Representations via Capacity Scaling, Distillation, and Low-Precision Inference

Model ReleasesDGX agent

arXiv:2604.09220v1 Announce Type: new Abstract: Implicit neural video representations encode entire video sequences within the parameters of a neural network and enable constant time frame reconstruct

Towards Lifelong Aerial Autonomy: Geometric Memory Management for Continual Visual Place Recognition in Dynamic Environments

Model ReleasesDGX agent

arXiv:2604.09038v1 Announce Type: cross Abstract: Robust geo-localization in changing environmental conditions is critical for long-term aerial autonomy. While visual place recognition (VPR) models pe

Unified Multimodal Uncertain Inference

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.08701v1 Announce Type: new Abstract: We introduce Unified Multimodal Uncertain Inference (UMUI), a multimodal inference task spanning text, audio, and video, where models must produce calib

V-CAGE: Vision-Closed-Loop Agentic Generation Engine for Robotic Manipulation

AgentsDGX agent

arXiv:2604.09036v1 Announce Type: new Abstract: Scaling Vision-Language-Action (VLA) models requires massive datasets that are both semantically coherent and physically feasible. However, existing sce

VAG: Dual-Stream Video-Action Generation for Embodied Data Synthesis

SafetyDGX agent

arXiv:2604.09330v1 Announce Type: cross Abstract: Recent advances in robot foundation models trained on large-scale human teleoperation data have enabled robots to perform increasingly complex real-wo

VisionFoundry: Teaching VLMs Visual Perception with Synthetic Images

ResearchDGX agent

arXiv:2604.09531v1 Announce Type: cross Abstract: Vision-language models (VLMs) still struggle with visual perception tasks such as spatial understanding and viewpoint recognition. One plausible contr

Watt Counts: Energy-Aware Benchmark for Sustainable LLM Inference on Heterogeneous GPU Architectures

Model ReleasesDGX agent

arXiv:2604.09048v1 Announce Type: cross Abstract: While the large energy consumption of Large Language Models (LLMs) is recognized by the community, system operators lack guidance for energy-efficient

we used Chandra-OCR-2 by @datalabto: https://huggingface.co/datalab-to/chandra-ocr-2 Full write-up by @NielsRogge: https://huggingface.co/bl…

IndustryDGX agent

Chandra-OCR-2 is an optical character recognition model developed by DataLab, available on Hugging Face at datalab-to/chandra-ocr-2. The model was highlighted by Hugging Face CEO Clément Delangue, wit

When Identity Skews Debate: Anonymization for Bias-Reduced Multi-Agent Reasoning

Model ReleasesDGX agent

arXiv:2510.07517v5 Announce Type: replace Abstract: Multi-agent debate (MAD) aims to improve large language model (LLM) reasoning by letting multiple agents exchange answers and then aggregate their o

12 Apr 2026

Can't get a good coding setup on Macbook Pro M3 Max 36GB

Local AiDGX agent

This Reddit thread from r/ollama discusses a user's difficulty achieving a satisfactory local AI coding assistant setup using Ollama on a MacBook Pro M3 Max with 36GB of unified memory. The discussion

Gemma 4 audio with MLX

Model ReleasesDGX agent

Thanks to a tip from Rahim Nathwani, here's a uv run recipe for transcribing an audio file on macOS using the 10.28 GB Gemma 4 E2B model with MLX and mlx-vlm: uv run --python 3.13 --with mlx_vlm --wit

Greg Rutkowski Anima Lora from Circlestone Labs (Anima makers) with training params

Local AiDGX agent

This Reddit post discusses a LoRA trained to emulate the style of digital artist Greg Rutkowski, built on top of Circlestone Labs' Anima model — a 2 billion parameter text-to-image model created via a

Most people building AI agents don't realize this until it's too late. Your agent's memory isn't a feature. It's your moat. 1. Closed harnes…

Model ReleasesDGX agent

Most people building AI agents don't realize this until it's too late. Your agent's memory isn't a feature. It's your moat. 1. Closed harness = they own your memory 2. Switch models → lose all context

Tile upscale controlnet with Z-Image-Base? Has anybody achieved good results?

Local AiDGX agent

This Reddit thread from r/StableDiffusion discusses community experiences using ControlNet Tile upscaling in combination with Z-Image-Base, a Stable Diffusion base model. ControlNet Tile models are of

Topping the charts!

AgentsDGX agent

Nous Research shared a post celebrating a model or benchmark achievement reaching the top of a performance leaderboard, likely referencing one of their Hermes or other open-source model releases. The

Z-Image Turbo Checkpoint - Deedeemegadoodo Edition

Local AiDGX agent

The 'Z-Image Turbo Checkpoint - Deedeemegadoodo Edition' is a community-shared checkpoint on r/StableDiffusion based on Z-Image Turbo, a distilled version of Z-Image, a 6B image model developed by the

11 Apr 2026

fine-tune LTX 2.3 with his own dataset?

Local AiDGX agent

This r/StableDiffusion thread discusses how to fine-tune the LTX-Video 2.3 model on a personal dataset, with the primary approach being LoRA (Low-Rank Adaptation), which fine-tunes a large AI model on

I got trolled

Local AiDGX agent

A Reddit post from the r/StableDiffusion community in which a user shares an experience of being trolled, likely related to AI image generation workflows, model recommendations, or settings advice. Th

I reduced my token usage by 178x in Claude Code!!

Model ReleasesDGX agent

A Reddit post from r/ollama describing how a user dramatically reduced their Claude Code token consumption by 178x, likely by routing simpler or lower-stakes tasks to a locally-run model via Ollama in

SDXL workflow

Local AiDGX agent

This Reddit post from r/StableDiffusion discusses an SDXL (Stable Diffusion XL) image generation workflow, likely covering pipeline setup, prompting strategies, and tool configurations using interface

10 Apr 2026

A Novel Automatic Framework for Speaker Drift Detection in Synthesized Speech

Model ReleasesDGX agent

arXiv:2604.06327v1 Announce Type: cross Abstract: Recent diffusion-based text-to-speech (TTS) models achieve high naturalness and expressiveness, yet often suffer from speaker drift, a subtle, gradual

AdaSpark: Adaptive Sparsity for Efficient Long-Video Understanding

ResearchDGX agent

arXiv:2604.08077v1 Announce Type: new Abstract: Processing long-form videos with Video Large Language Models (Video-LLMs) is computationally prohibitive. Current efficiency methods often compromise fi

Advanced inpaint/edit Klein/Qwen workflows

Model ReleasesDGX agent

A Reddit post on r/StableDiffusion discussing advanced ComfyUI workflows that combine the FLUX Klein and Qwen Image Edit models for precision inpainting and image editing tasks. FLUX Klein offers ...

AgentGate: A Lightweight Structured Routing Engine for the Internet of Agents

Model ReleasesDGX agent

arXiv:2604.06696v1 Announce Type: new Abstract: The rapid development of AI agent systems is leading to an emerging Internet of Agents, where specialized agents operate across local devices, edge node

AnchorSplat: Feed-Forward 3D Gaussian Splatting with 3D Geometric Priors

Model ReleasesDGX agent

arXiv:2604.07053v2 Announce Type: replace Abstract: Recent feed-forward Gaussian reconstruction models adopt a pixel-aligned formulation that maps each 2D pixel to a 3D Gaussian, entangling Gaussian r

Anthropic and OpenAI target big businesses with enterprise-grade controls and lower pricing

Model ReleasesDGX agent

Artificial intelligence leaders Anthropic PBC and OpenAI Group PBC are stepping up their efforts to compete for the enterprise, making their most advanced agentic tools more accessible to the largest

Anticipating tipping in spatiotemporal systems with machine learning

Model ReleasesDGX agent

arXiv:2604.06454v1 Announce Type: cross Abstract: In nonlinear dynamical systems, tipping refers to a critical transition from one steady state to another, typically catastrophic, steady state, often

Are GUI Agents Focused Enough? Automated Distraction via Semantic-level UI Element Injection

SafetyDGX agent

arXiv:2604.07831v1 Announce Type: cross Abstract: Existing red-teaming studies on GUI agents have important limitations. Adversarial perturbations typically require white-box access, which is unavaila

Bad news on Happy Horse from twitter

Local AiDGX agent

HappyHorse-1.0 is a pseudonymous AI video generation model that appeared on April 7, 2026, topping the Artificial Analysis Video Arena leaderboard in both text-to-video and image-to-video (no audio...

Behind the Analysis with Google Cloud and Team USA: Architecting AI infrastructure for U.S. Winter Olympians

HardwareDGX agent

In freeskiing and snowboarding, traditional video replay shows you what happened during a complex aerial maneuver, but it fails to explain the physics of how it was possible. At the speed of the sport

Benchmarking LLM Tool-Use in the Wild

Model ReleasesDGX agent

arXiv:2604.06185v1 Announce Type: cross Abstract: Fulfilling user needs through Large Language Model multi-turn, multi-step tool-use is rarely a straightforward process. Real user interactions are inh

Broken by Default: A Formal Verification Study of Security Vulnerabilities in AI-Generated Code

Model ReleasesDGX agent

arXiv:2604.05292v2 Announce Type: replace-cross Abstract: AI coding assistants are now used to generate production code in security-sensitive domains, yet the exploitability of their outputs remains u

Continual Visual Anomaly Detection on the Edge: Benchmark and Efficient Solutions

Model ReleasesDGX agent

arXiv:2604.06435v1 Announce Type: cross Abstract: Visual Anomaly Detection (VAD) is a critical task for many applications including industrial inspection and healthcare. While VAD has been extensively

Countering the Over-Reliance Trap: Mitigating Object Hallucination for LVLMs via a Self-Validation Framework

ResearchDGX agent

arXiv:2601.22451v2 Announce Type: replace-cross Abstract: Despite progress in Large Vision Language Models (LVLMs), object hallucination remains a critical issue in image captioning task, where models

DeCo: Frequency-Decoupled Pixel Diffusion for End-to-End Image Generation

ResearchDGX agent

arXiv:2511.19365v2 Announce Type: replace-cross Abstract: Pixel diffusion aims to generate images directly in pixel space in an end-to-end fashion. This approach avoids the limitations of VAE in the t

Differentially Private Language Generation and Identification in the Limit

ResearchDGX agent

arXiv:2604.08504v1 Announce Type: cross Abstract: We initiate the study of language generation in the limit, a model recently introduced by Kleinberg and Mullainathan [KM24], under the constraint of d

Efficient Quantization of Mixture-of-Experts with Theoretical Generalization Guarantees

ResearchDGX agent

arXiv:2604.06515v1 Announce Type: cross Abstract: Sparse Mixture-of-Experts (MoE) allows scaling of language and vision models efficiently by activating only a small subset of experts per input. While

Evaluating In-Context Translation with Synchronous Context-Free Grammar Transduction

ResearchDGX agent

arXiv:2604.07320v1 Announce Type: cross Abstract: Low-resource languages pose a challenge for machine translation with large language models (LLMs), which require large amounts of training data. One p

Evaluating Repository-level Software Documentation via Question Answering and Feature-Driven Development

Model ReleasesDGX agent

arXiv:2604.06793v1 Announce Type: cross Abstract: Software documentation is crucial for repository comprehension. While Large Language Models (LLMs) advance documentation generation from code snippets

EvoFlows: Evolutionary Edit-Based Flow-Matching for Protein Engineering

ResearchDGX agent

arXiv:2603.11703v2 Announce Type: replace Abstract: We introduce EvoFlows, a variable-length protein sequence-to-sequence modeling approach designed for protein engineering. Existing protein language

Fighting AI with AI: AI-Agent Augmented DNS Blocking of LLM Services during Student Evaluations

Model ReleasesDGX agent

arXiv:2604.02360v1 Announce Type: cross Abstract: The transformative potential of large language models (LLMs) in education, such as improving accessibility and personalized learning, is being eclipse

glm 5.1 is doing well

Local AiDGX agent

GLM-5.1 is Z.ai's next-generation flagship model for agentic engineering, built on a 754-billion parameter Mixture-of-Experts architecture with 40 billion active parameters per token, a 200,000-tok...

How Much LLM Does a Self-Revising Agent Actually Need?

AgentsDGX agent

arXiv:2604.07236v2 Announce Type: new Abstract: Recent LLM-based agents often place world modeling, planning, and reflection inside a single language model loop. This can produce capable behavior, but

Hybrid ResNet-1D-BiGRU with Multi-Head Attention for Cyberattack Detection in Industrial IoT Environments

ResearchDGX agent

arXiv:2604.06481v1 Announce Type: cross Abstract: This study introduces a hybrid deep learning model for intrusion detection in Industrial IoT (IIoT) systems, combining ResNet-1D, BiGRU, and Multi-Hea

In a world where writing code to build websites and apps is trivial (thank you Lovable, Cursor, Claude,...), the real differentiation for yo…

Model ReleasesDGX agent

In a world where writing code to build websites and apps is trivial (thank you Lovable, Cursor, Claude,...), the real differentiation for you and your company (and what makes you successful) will be h

InstAP: Instance-Aware Vision-Language Pre-Train for Spatial-Temporal Understanding

Model ReleasesDGX agent

arXiv:2604.08337v1 Announce Type: new Abstract: Current vision-language pre-training (VLP) paradigms excel at global scene understanding but struggle with instance-level reasoning due to global-only s

k-Maximum Inner Product Attention for Graph Transformers and the Expressive Power of GraphGPS

Model ReleasesDGX agent

arXiv:2604.03815v2 Announce Type: replace-cross Abstract: Graph transformers have shown promise in overcoming limitations of traditional graph neural networks, such as oversquashing and difficulties i

Kuramoto Oscillatory Phase Encoding: Neuro-inspired Synchronization for Improved Learning Efficiency

Model ReleasesDGX agent

arXiv:2604.07904v1 Announce Type: cross Abstract: Spatiotemporal neural dynamics and oscillatory synchronization are widely implicated in biological information processing and have been hypothesized t

Learning the Stellar Structure Equations via Self-supervised Physics-Informed Neural Networks

Model ReleasesDGX agent

arXiv:2604.06255v1 Announce Type: cross Abstract: Stellar astrophysics relies critically on accurate descriptions of the physical conditions inside stars. Traditional solvers such as exttt{MESA} (Mo

LLM Prompt Duel Optimizer: Efficient Label-Free Prompt Optimization

ResearchDGX agent

arXiv:2510.13907v3 Announce Type: replace Abstract: Large language models (LLMs) are highly sensitive to prompts, but most automatic prompt optimization (APO) methods assume access to ground-truth ref

Logics-Parsing-Omni Technical Report

Model ReleasesDGX agent

arXiv:2603.09677v3 Announce Type: replace Abstract: Addressing the challenges of fragmented task definitions and the heterogeneity of unstructured data in multimodal parsing, this paper proposes the O

Making Room for AI: Multi-GPU Molecular Dynamics with Deep Potentials in GROMACS

Model ReleasesDGX agent

arXiv:2604.07276v1 Announce Type: cross Abstract: GROMACS is a de-facto standard for classical Molecular Dynamics (MD). The rise of AI-driven interatomic potentials that pursue near-quantum accuracy a

MedRoute: RL-Based Dynamic Specialist Routing in Multi-Agent Medical Diagnosis

AgentsDGX agent

arXiv:2604.06180v1 Announce Type: cross Abstract: Medical diagnosis using Large Multimodal Models (LMMs) has gained increasing attention due to capability of these models in providing precise diagnose

MedVR: Annotation-Free Medical Visual Reasoning via Agentic Reinforcement Learning

SafetyDGX agent

arXiv:2604.08203v1 Announce Type: new Abstract: Medical Vision-Language Models (VLMs) hold immense promise for complex clinical tasks, but their reasoning capabilities are often constrained by text-on

MICA: Multivariate Infini Compressive Attention for Time Series Forecasting

ResearchDGX agent

arXiv:2604.06473v1 Announce Type: new Abstract: Multivariate forecasting with Transformers faces a core scalability challenge: modeling cross-channel dependencies via attention compounds attention's q

Mitigating Domain Drift in Multi Species Segmentation with DINOv2: A Cross-Domain Evaluation in Herbicide Research Trials

ApplicationsDGX agent

arXiv:2508.07514v3 Announce Type: replace Abstract: Reliable plant species and damage segmentation for herbicide field research trials requires models that can withstand substantial real-world variati

NaviSplit: Dynamic Multi-Branch Split DNNs for Efficient Distributed Autonomous Navigation

AgentsDGX agent

arXiv:2406.13086v2 Announce Type: replace Abstract: Lightweight autonomous unmanned aerial vehicles (UAV) are emerging as a central component of a broad range of applications. However, autonomous navi

Neural Computers

SafetyDGX agent

arXiv:2604.06425v1 Announce Type: cross Abstract: We propose a new frontier: Neural Computers (NCs) -- an emerging machine form that unifies computation, memory, and I/O in a learned runtime state. Un

Neural parametric representations for thin-shell shape optimisation

Model ReleasesDGX agent

arXiv:2604.06612v1 Announce Type: cross Abstract: Shape optimisation of thin-shell structures requires a flexible, differentiable geometric representation suitable for gradient-based optimisation. We

Physics-Informed Neural Networks for Joint Source and Parameter Estimation in Advection-Diffusion Equations

Model ReleasesDGX agent

arXiv:2512.07755v3 Announce Type: replace-cross Abstract: Recent studies have demonstrated the success of deep learning in solving forward and inverse problems in engineering and scientific computing

← Previous
1…371372373374375…1050
Next →