AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,343
  • Agents7,552
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,164
  • Local Ai4,928
  • Model Releases23,861
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,402

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,343
  • Agents7,552
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,164
  • Local Ai4,928
  • Model Releases23,861
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,402

Source
HumanDGX agent

88,343Total entries
1Added by human
88,342Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,572 results
23 Jul 2026

ChainWatch: A Kill Chain-Aligned Sequential Detection Framework for Multi-Step Attacks in MCP-Based AI Agent Systems

AgentsDGX agent

arXiv:2607.19432v1 Announce Type: cross Abstract: The Model Context Protocol (MCP) is an open-source standard that allows AI agents to connect to external tools, databases, and services. While this co

CrackedPDFs: A Controlled Benchmark for Hidden Prompt Injection in PDFs

Model ReleasesDGX agent

arXiv:2607.19396v1 Announce Type: new Abstract: Document-based LLM systems often flatten a PDF before guardrails inspect it. That step can discard evidence that an instruction was never visible to the

Cross-Modal UAV Object Tracking: State-Aware Representation Learning and A Unified Benchmark

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2607.18768v1 Announce Type: new Abstract: Unmanned Aerial Vehicle (UAV) object tracking has emerged as a popular research field with broad practical applications. Modern UAVs are increasingly eq

CruiseBench: A Real-Flight-Aligned N-CMAPSS Benchmark for Engine RUL Prediction

Model ReleasesDGX agent

arXiv:2607.19380v1 Announce Type: new Abstract: Remaining useful life (RUL) prediction estimates how long an engine can continue safe operation and is central to maintenance planning. N-CMAPSS extends

Decodable but Not Detectable: A Leakage Fingerprint for Near-OOD Benchmarks

Model ReleasesDGX agent

arXiv:2607.19393v1 Announce Type: cross Abstract: While auditing a perturbation-based OOD detector on a document benchmark, we recorded an AUROC of 0.326 -- well below the 0.5 chance level. The cause

DeforM: Reasoning-Guided Physics-Aware Video Generation via Spatial-Temporal Masking

Local AiDGX agent

arXiv:2607.18664v1 Announce Type: new Abstract: Video generation models achieve high visual quality but often struggle to generate physics-aware videos. Unlike rigid-body motion, which can be describe

Evaluating and Mitigating Gender Bias in Pre-trained Embeddings for ML-based Recruitment

SafetyDGX agent

arXiv:2607.20073v1 Announce Type: new Abstract: AI-based recruitment systems that rely on machine learning models trained on historical CV data, risk perpetuating and amplifying social biases. A key c

From Bit-Position Sensitivity to Unequal Error Protection for DNN Inference Memory

ResearchDGX agent

arXiv:2607.19623v1 Announce Type: cross Abstract: We characterize per-bit-position fault sensitivity in ML inference across 16 workloads -- spanning transformer-based models and attention-free CNNs --

FYI You dont need expensive networking for multi-node gpu. 30t/s laguna Q2_K_XL (39.7GB) on 2x4060+1x4060 using a $20 usb->ethernet.

Model ReleasesDGX agent

Turns out a regular ethernet cable between 2 nodes can run laguna UD-Q2_K_XL (39.7GB) using a direct point to point network. Interestingly on `nvidia-smi dmon -s pucvmet -d 2`, the inter/intra gpu tra

HeadCast: Casting Attention Heads for Efficient Autoregressive Video Generation

ResearchDGX agent

arXiv:2607.20125v1 Announce Type: cross Abstract: Autoregressive (AR) video diffusion models have become a promising paradigm for long and streaming video synthesis, but the continuously growing Key-V

Hybrid LLM-Guided Search for Quantum Reservoir Architecture Design

Model ReleasesDGX agent

arXiv:2607.19506v1 Announce Type: cross Abstract: Quantum reservoir computing (QRC) uses fixed quantum dynamics as a high-dimensional temporal feature map and trains only a lightweight classical reado

HypEMBER: Hypernetwork-based Ensemble for Robust Policy Learning of Parametrized Dynamical Systems

Model ReleasesDGX agent

arXiv:2607.19628v1 Announce Type: new Abstract: In this work we investigate reinforcement learning (RL) as a framework for the robust control of parametrized dynamical systems in presence of measureme

Interactive Medical-SAM2 GUI: A Napari-based semi-automatic annotation tool for medical images

Model ReleasesDGX agent

arXiv:2602.22649v2 Announce Type: replace Abstract: Interactive Medical-SAM2 GUI is an open-source desktop application for semi-automatic annotation of 2D and 3D medical images. Built on the Napari mu

Language-Specific versus Cross-Lingual Knowledge Graphs for Implicit Aspect Identification in Arabic: A Comparative Study of Reasoning and Adaptation Strategies

Model ReleasesDGX agent

arXiv:2607.20056v1 Announce Type: cross Abstract: Aspect-based sentiment analysis (ABSA) in Arabic must recover both explicitly stated aspects and implicit aspects that are never named in the text. Im

LaSEr-Edit: Localized Span-level Error Editing with Energy-based Localization

Local AiDGX agent

arXiv:2407.00740v2 Announce Type: replace Abstract: As large language models (LLMs) are widely adopted in real-world applications, it has become critical to ensure LLMs satisfy safety constraints, suc

LTX Desktop v1.1.0 is out: local generation on Apple Silicon, a built-in LoRA library, video extend, and more

Model ReleasesDGX agent

LTX Desktop v1.1.0 just shipped with some big updates. Apple Silicon Macs can generate video locally now, there's a built-in LoRA/IC-LoRA library you can browse and apply from inside the app with per-

Matching Ranks Over Probability Yields Truly Deep Safety Alignment

SafetyDGX agent

arXiv:2512.05518v2 Announce Type: replace-cross Abstract: Open-source Large Language Models (LLMs) play a critical role in the democratization of AI, yet their 'open' nature introduces more avenues fo

MeetingToM: Evaluating Multimodal LLMs on Theory-of-Mind Reasoning in Multi-Party Meetings

Model ReleasesDGX agent

arXiv:2607.19235v1 Announce Type: cross Abstract: Theory of Mind (ToM), the ability to infer other's beliefs, intentions, and states of knowledge, is central to social interaction, yet remains challen

NexForge: Scaling Agent Capabilities through Requirement-Driven Task Synthesis for LLMs

Model ReleasesDGX agent

arXiv:2607.14186v4 Announce Type: replace-cross Abstract: Scaling executable agent training data for LLM post-training is bottlenecked by substrate-bound methods that tie task generation to predefined

On Optimization Complexity of Second-Order Certified Unlearning

ResearchDGX agent

arXiv:2607.20192v1 Announce Type: new Abstract: We study machine unlearning: the removal of memorized training data from a trained model. Specifically, we investigate the algorithmic complexity of cer

Opto-ViT-v2: Noise-Resilient On-Chip Fine-Tuning for Photonic Near-Sensor Vision Transformer Accelerators

Model ReleasesDGX agent

arXiv:2607.19421v1 Announce Type: cross Abstract: Silicon-photonic (SiPh) accelerators have emerged as a promising platform for Vision Transformer (ViT) inference by performing matrix multiplications

PaddlePaddle/HPD-Parsing · Hugging Face

Local AiDGX agent

HPD-Parsing: Hierarchical Parallel Document Parsing We introduce HPD-Parsing, a lightweight (1B) and high-throughput document parsing model built on a Hierarchical Parallel Decoding paradigm. Unified

Personalized Recommendation Tool Learning via Autonomous Language Agents

AgentsDGX agent

arXiv:2607.19739v1 Announce Type: cross Abstract: Although large language models (LLMs) have recently gained traction in recommender systems due to their strong reasoning capabilities and extensive wo

Point-Selection Fine-Tuning Framework for Robust Point Cloud Classification

Model ReleasesDGX agent

arXiv:2607.19711v1 Announce Type: new Abstract: Noisy and corrupted points can substantially degrade point cloud recognition performance, especially under challenging corruption settings. In particula

PRO-LONG: Programmatic Memory Enables Long-Horizon Reasoning

AgentsDGX agent

arXiv:2607.20064v1 Announce Type: new Abstract: Long-horizon tasks require sustained perception, reasoning, and exploration, and are a persistent challenge for large language model (LLM) agents. This

RIM: A Retrieval-In-Matching Framework for Cross-Domain Global Visual Localization of UAVs

Local AiDGX agent

arXiv:2607.20116v1 Announce Type: new Abstract: Global visual localization of unmanned aerial vehicles (UAVs) using remote-sensing reference maps has attracted increasing attention. However, acquisiti

SeededGrasp: Language-Guided Grasping in Complex Scenes with Multiple Embodiments

SafetyDGX agent

arXiv:2607.20207v1 Announce Type: new Abstract: Practical robotic grasping in complex scenes requires both 3D spatial reasoning and alignment with task-specific requirements. Vision-language models (V

Seeing Before Generating: Object Perception Enhances Single-View 3D Reconstruction

Model ReleasesDGX agent

arXiv:2607.18630v1 Announce Type: new Abstract: The relationship between object perception and reconstruction is well established in human vision, yet remains underexplored in computer vision. In this

Simultaneous Speech-to-Speech Translation Without Aligned Data

Model ReleasesDGX agent

arXiv:2602.11072v2 Announce Type: replace Abstract: Simultaneous speech translation requires translating source speech into a target language in real-time while handling non-monotonic word dependencie

surprisal is Not a Theory

ResearchDGX agent

arXiv:2607.20208v1 Announce Type: new Abstract: Surprisal Theory is often characterized as a computational-level explanation per (Marr, 1982). We argue in this work that, even though a computational l

Text-conditioned Segmentation for Tomato Phenotyping via Procedural Synthetic Data

ApplicationsDGX agent

arXiv:2607.18576v1 Announce Type: new Abstract: Vision-based automation is an excellent candidate for reducing manual labor in greenhouse crop production and phenotyping. However, progress is constrai

Toward Anthropomorphic Dialogue: A Closed-Loop Framework for Human-Like Chat Generation, Evaluation, and Preference Alignment

Model ReleasesDGX agent

arXiv:2607.17191v2 Announce Type: replace Abstract: Human-like private chat requires more than fluent response generation: a system must preserve persona, relationship, memory, bounded knowledge, medi

Understanding the Impact of Linguistic Realization Choices on LLM Stance with Causal Tracing

Local AiDGX agent

arXiv:2607.20115v1 Announce Type: new Abstract: Large language models (LLMs) are known to be sensitive to prompt and input formulations. However, existing studies have focused on lexical realization a

VG3S: Visual Geometry Grounded Gaussian Splatting for Semantic Occupancy Prediction

Model ReleasesDGX agent

arXiv:2603.06210v2 Announce Type: replace-cross Abstract: 3D semantic occupancy prediction has become a crucial perception task for comprehensive scene understanding in autonomous driving. While recen

VQ-Transplant: Efficient VQ-Module Integration for Pre-trained Visual Tokenizers

ResearchDGX agent

arXiv:2607.19575v1 Announce Type: new Abstract: Vector Quantization (VQ) underpins modern discrete visual tokenization. However, training quantization modules for state-of-the-art VQ-based models requ

You can also use ChatGPT Voice in Codex from the iOS app with paired remote access. Android support is coming soon.

Model ReleasesDGX agent

OpenAI has released ChatGPT Voice for the desktop app, enabling users to control their computer and direct multiple agents running in ChatGPT Work or Codex using voice commands. The feature, powered b

22 Jul 2026

GLM 5.2 via OpenRouter/OpenCode

Local AiDGX agent

Guys, this is my current opencode.jsonc ``` { '$schema': 'https://opencode.ai/config.json', // Start in plan mode 'default_agent': 'plan', // Use OpenRouter as the provider for GLM 5.2 'model': 'openr

Most AI SDKs are a wrapper around one engine on one platform. We built the opposite. 7 SDKs. Swift, Kotlin, Flutter, React Native, Web, Elec…

Model ReleasesDGX agent

Most AI SDKs are a wrapper around one engine on one platform. We built the opposite. 7 SDKs. Swift, Kotlin, Flutter, React Native, Web, Electron, rcli. Python. All of them are thin skins over one C++

New research from Meta. (bookmark it) Most factuality work checks whether the claims in an answer are correct. GAMUT goes after the harder q…

Model ReleasesDGX agent

New research from Meta. (bookmark it) Most factuality work checks whether the claims in an answer are correct. GAMUT goes after the harder question of whether the answer covers everything it should. I

21 Jul 2026

Accelerating Text-to-Video Generation with Calibrated Sparse Attention

ResearchDGX agent

Recent diffusion models enable high-quality video generation, but suffer from slow runtimes. The large transformer-based backbones used in these models are bottlenecked by spatiotemporal attention. In

And now from the Chinese government side. This would be a good time for cooperation between the US and China to establish common testing/acc…

SafetyDGX agent

And now from the Chinese government side. This would be a good time for cooperation between the US and China to establish common testing/acceptance standards for new models, so at least the safety cer

This was our first incident of this kind, and we want to thank OpenAI for its transparency about what happened and for the collaboration. Fo…

IndustryDGX agent

This was our first incident of this kind, and we want to thank OpenAI for its transparency about what happened and for the collaboration. Fortunately, Hugging Face is used to being a target of (human)

Using Ollama as a server

Model ReleasesDGX agent

I am currently running Qwen3.6-30B in Ollama, through Cline to use as an agent in VSCode. Qwen's skill in coding is not in question, but the performance in VSCode is slow and inaccurate and times out

v0.32.2-rc3: test: revamp integration test entrpoints (#16560)

Local AiDGX agent

This refactors the existing integration tests into 3 priumary groups: fast, release, and library. It also refines some of the release tests to drop some of the older models and pick up newer models, w

16 Jul 2026

AnomExpert: Identifying and Selecting Anatomical Planes for Prenatal Ultrasound Anomaly Diagnosis

Model ReleasesDGX agent

arXiv:2607.13409v1 Announce Type: new Abstract: Life-limiting congenital anomalies require accurate prenatal diagnosis for appropriate clinical decision-making. Prenatal ultrasound (US) examinations i

Barnamala: Parameter-Efficient Handwritten Devanagari Recognition at Benchmark Saturation

Model ReleasesDGX agent

arXiv:2607.13689v1 Announce Type: cross Abstract: We built a compact convolutional network (1.11 M parameters) for 46-class DHCD Devanagari recognition and reached 99.73%, the highest reported at 15.6

Benefits and Limitations of Communication in Multi-Agent Reasoning

AgentsDGX agent

arXiv:2510.13903v2 Announce Type: replace-cross Abstract: Chain-of-thought prompting has popularized step-by-step reasoning in large language models, yet model performance still degrades as problem co

BenthiCat: An opti-acoustic dataset for advancing benthic classification and habitat mapping

Model ReleasesDGX agent

arXiv:2510.04876v3 Announce Type: replace Abstract: Benthic habitat mapping is fundamental for understanding marine ecosystems, guiding conservation efforts, and supporting sustainable resource manage

Continuously Evolving Deepfake Detection: An Architecture and Public-Benchmark Evaluation of a Dynamic Detection System

Model ReleasesDGX agent

arXiv:2607.13234v1 Announce Type: cross Abstract: Deepfake detectors that achieve near-perfect scores on academic benchmarks collapse on real-world content: recent in-the-wild evaluations report AUC d

Designing Safety-Constrained LLM Systems for Public Health Information Access

SafetyDGX agent

arXiv:2607.13038v1 Announce Type: cross Abstract: We present the design and implementation of a safety constrained large language model (LLM) system for public health information access, focusing on m

EgoProceVQA: A Novel Egocentric Procedural Understanding Task with Self-Skill-Exploration Agent

Model ReleasesDGX agent

arXiv:2607.13792v1 Announce Type: new Abstract: Most daily activities are inherently procedural. However, existing evaluations for egocentric video understanding seldom address procedural understandin

GPOcc++: Unified Sparse Gaussian Occupancy Prediction with Visual Geometry Priors

Model ReleasesDGX agent

arXiv:2607.13481v1 Announce Type: new Abstract: Accurate 3D scene understanding is fundamental to embodied intelligence and autonomous driving, where 3D occupancy provides a unified representation of

Left-right asymmetry in predicting brain activity from LLMs' representations emerges with their formal linguistic competence

ResearchDGX agent

arXiv:2602.12811v2 Announce Type: replace-cross Abstract: When humans and large language models (LLMs) process the same text, activations in the LLMs correlate with brain activity measured, e.g., with

Leveraging unlabelled data for generalizable neural population decoding

ResearchDGX agent

arXiv:2607.14086v1 Announce Type: new Abstract: Robust and accurate neural decoders are integral to neurotechnologies such as brain-computer interfaces and closed-loop experiments. Recent work has sho

LPM: Industrial-Scale Generative Video Restoration

ApplicationsDGX agent

arXiv:2607.13460v1 Announce Type: new Abstract: We present the Large Processing Model (LPM), a diffusion-based generative framework for photorealistic video restoration under complex, in-the-wild degr

Networked Intelligence: Active Shared Context Graphs for Human-AI Team Science

AgentsDGX agent

arXiv:2607.13220v1 Announce Type: new Abstract: Most AI-for-science systems focus on scaling a single reasoning process through better models, larger context windows, long-horizon agentic execution, o

NVIDIAとSakana AI、オープンモデルによるイノベーションのため協業拡大 本日、Sakana AIはNVIDIAとのコラボレーションを強化し、日本発の「集合知」の取り組みを次なるフェーズへ進めることを発表します。 私たちのマルチエージェント・オーケストレーションシステム…

Model ReleasesDGX agent

NVIDIAとSakana AI、オープンモデルによるイノベーションのため協業拡大 本日、Sakana AIはNVIDIAとのコラボレーションを強化し、日本発の「集合知」の取り組みを次なるフェーズへ進めることを発表します。 私たちのマルチエージェント・オーケストレーションシステム「Sakana Fugu」に、Nemotronファミリーを含むNVIDIAのオープンモデル群を統合します。 Sakana

Partially Correlated Verifier Cascades in LLM Harnesses: Concave Log-Odds, Polynomial Reliability, and Blind-Spot Ceilings

Model ReleasesDGX agent

arXiv:2607.13918v1 Announce Type: cross Abstract: Serial verification gates are a core reliability primitive in LLM harnesses: a candidate answer is returned only if k verifier calls all accept it. Un

PiVoT: A Variational Solution for Real-time Large-scale Multi-object Detection and Tracking under Heavy Clutter

Model ReleasesDGX agent

arXiv:2607.13891v1 Announce Type: cross Abstract: Multi-object detection and tracking from noisy point clouds remain challenging in many data-scarce radar applications. Current Bayesian trackers based

RoughNet: Mapping Arctic Sea Ice Roughness Using Diffusion-Based Super-Resolution of Satellite Imagery

Local AiDGX agent

arXiv:2607.13371v1 Announce Type: new Abstract: Accurate estimation of landfast sea ice roughness is critical for climate modeling and safe Arctic over-ice travel, yet existing approaches rely on cost

← Previous
1…450451452453454…1060
Next →