AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries92,405
  • Agents7,865
  • Applications5,605
  • Concepts5
  • Hardware1,963
  • Industry6,239
  • Local Ai5,175
  • Model Releases25,270
  • Research21,121
  • Safety13,951
  • Syntheses17
  • Tools1,680
  • Tutorials3,514

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries92,405
  • Agents7,865
  • Applications5,605
  • Concepts5
  • Hardware1,963
  • Industry6,239
  • Local Ai5,175
  • Model Releases25,270
  • Research21,121
  • Safety13,951
  • Syntheses17
  • Tools1,680
  • Tutorials3,514

Source
HumanDGX agent

Content type
92,405Total entries
1Added by human
92,404Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
66,927 results
Model Releases

MAVISEG: Manifold Propagation and Visual Prototypes for Zero-Shot Open-Vocabulary Segmentation in Diffusion Transformers

DGX agent

arXiv:2608.05878v1 Announce Type: new Abstract: Text-to-image diffusion transformers learn about objects and scenes by learning to generate them, making them strong candidates for training-free zero-s

model-releasesarxiv-cs-cv
7 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Mind the Gaps: Mixture-of-Minds for Human Simulation

DGX agent

arXiv:2608.06115v1 Announce Type: new Abstract: Predicting how a population will answer a new question is a long-standing goal. Statistical methods succeed at the level of the mass but falter at the l

model-releasesarxiv-cs-ai
7 Aug 2026
Research

MOSAIK: Multi-Patch Content-Aware Spatial Allocation of Image Tokens for Efficient Generation

DGX agent

arXiv:2608.05450v1 Announce Type: new Abstract: Pixel-space diffusion models avoid the reconstruction ceiling of latent diffusion models by generating directly in image space. However, their substanti

researcharxiv-cs-cv
7 Aug 2026
Model Releases

MS-MLB: An Open Machine Learning Benchmark for Blood-Based MS Classification

DGX agent

arXiv:2608.05196v1 Announce Type: new Abstract: Multiple sclerosis (MS) is diagnosed through clinical assessment, magnetic resonance imaging, laboratory evidence when appropriate, and exclusion of bet

model-releasesarxiv-cs-lg
7 Aug 2026
Model Releases

NavTrust: Benchmarking Trustworthiness for Embodied Navigation

DGX agent

arXiv:2603.19229v2 Announce Type: replace-cross Abstract: There are two major categories of embodied navigation: Vision-Language Navigation (VLN), where agents navigate by following natural language i

model-releasesarxiv-cs-ai
7 Aug 2026
Research

Once a Response, Always a Response: Detecting LLM-generated Text via Latent Prompt Restoration

DGX agent

arXiv:2608.05741v1 Announce Type: cross Abstract: Large language models (LLMs) can generate fluent and convincing text at scale, creating growing risks for misinformation dissemination, educational mi

researcharxiv-cs-ai
7 Aug 2026
Research

Online Reasoning Calibration: Test-Time Training Enables Generalizable Conformal LLM Reasoning

DGX agent

arXiv:2604.01170v2 Announce Type: replace-cross Abstract: While test-time scaling has enabled large language models to solve highly difficult tasks, state-of-the-art results come at exorbitant compute

researcharxiv-cs-ai
7 Aug 2026
Research

Pixel-TTS: Image based Text Rendering for Robust Text-to-Speech

DGX agent

arXiv:2606.14750v2 Announce Type: replace-cross Abstract: Recent advances in pixel-based text modeling show that representing text as images enables models to exploit visual cues for language understa

researcharxiv-cs-ai
7 Aug 2026
Model Releases

Recursive Synthesis for Long-Horizon Terminal Tasks

DGX agent

arXiv:2608.05466v1 Announce Type: new Abstract: High-quality long-horizon training data for terminal agents is expensive to produce, often costing hundreds to thousands of dollars per task, because ea

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

StepReflect: Structured UI Transition Reflection for Mobile GUI Agents

DGX agent

arXiv:2608.05587v1 Announce Type: new Abstract: Autonomous mobile GUI agents require accurate action reflection for reliable long-horizon execution. Existing approaches rely on open-ended multimodal r

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

The backgrounds were generated with MiniMax M3. Everything else, including the full game design, was done with DeepSeek Flash. Most importan…

DGX agent

The backgrounds were generated with MiniMax M3. Everything else, including the full game design, was done with DeepSeek Flash. Most importantly, all of this was done inside Hermes Agent. You won't bel

model-releasesnous-research--x
7 Aug 2026
Model Releases

The Next Screenshot Knows: Gated Hindsight Distillation for Mobile GUI Agents

DGX agent

arXiv:2608.06065v1 Announce Type: new Abstract: GUI agents are commonly trained offline from successful interaction trajectories. Standard training decomposes each trajectory into prefix-action pairs:

model-releasesarxiv-cs-cv
7 Aug 2026
Research

Trust-Based Incentive Mechanisms in Semi-Decentralized Federated Learning Systems

DGX agent

arXiv:2602.08290v2 Announce Type: replace-cross Abstract: In federated learning (FL), decentralized model training allows multi-ple participants to collaboratively improve a shared machine learning mo

researcharxiv-cs-ai
7 Aug 2026
Research

TruthLens: Object Hallucination Detection via Self-Evaluating Truthfulness Scores in LVLMs

DGX agent

arXiv:2608.05616v1 Announce Type: new Abstract: Despite the remarkable progress of large vision language models (LVLMs), object hallucination remains a fundamental challenge that hinders their trustwo

researcharxiv-cs-cv
7 Aug 2026
Model Releases

Vibe Compiler: A Research-Logic Synthesis Tool That Runs without Prompt Engineering -Toward Enhancing Metacognition for Sustaining Agency in the Age of Generative AI-

DGX agent

arXiv:2608.05545v1 Announce Type: cross Abstract: Generative AI used as a capable servant has greatly accelerated intellectual work, but it also risks eroding human epistemic agency by encouraging unc

model-releasesarxiv-cs-ai
7 Aug 2026
Research

Vorch-Omni: Multi-Task Orchestration of Sight and Sound

DGX agent

arXiv:2608.05803v1 Announce Type: new Abstract: Recent advances in generative video modeling have enabled diverse generation, reference-based synthesis, extension, and editing, but existing approaches

researcharxiv-cs-cv
7 Aug 2026
Local Ai

When Does Consensus Mean Correctness? Measuring the Agreement-Accuracy Coupling with Semantics-Preserving Re-Rendering

DGX agent

arXiv:2608.05670v1 Announce Type: new Abstract: A model's agreement across perturbed inputs is used both as a label-free reliability signal and as a self-training target, on the premise that agreement

local-aiarxiv-cs-lg
7 Aug 2026
Model Releases

Adversarial Attacks for Good: A Survey of Proactive Protection across the Visual Content Lifecycle

DGX agent

arXiv:2608.04314v1 Announce Type: cross Abstract: Once visual content enters an AI pipeline, its owner often retains little technical control over how it is used. Legal and regulatory remedies can add

model-releasesarxiv-cs-cv
6 Aug 2026
Safety

Among proponents of neurosymbolic architectures, there had been some debate over the years about whether the outer level would be symbolic (…

DGX agent

Among proponents of neurosymbolic architectures, there had been some debate over the years about whether the outer level would be symbolic (i.e. a harness that calls neural models) or whether the oute

safetygary-marcus--x
6 Aug 2026
Model Releases

An Update to Sir Shortoken: Introducing LELP-S+ (Less English, Less Prose)

DGX agent

A small update to Sir Shortoken. Sir Shortoken already had Quick, Balanced, Deep, Bullets, and Aggressive Bullets. I wanted something between Bullets and normal prose. So I added LELP-S+ (Less English

model-releasesr-ollama
6 Aug 2026
Model Releases

Argus: A General-Purpose Agentic Runtime for Long-Horizon Reasoning

DGX agent

arXiv:2608.05144v1 Announce Type: new Abstract: Long-horizon reasoning requires an agentic runtime that can persist when evidence supports its current approach and pivot when measurements reveal failu

model-releasesarxiv-cs-ai
6 Aug 2026
Safety

Beyond the Dirac Delta: Mitigating Diversity Collapse in Reinforcement Fine-Tuning for Versatile Image Generation

DGX agent

arXiv:2601.12401v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) has emerged as a powerful paradigm for fine-tuning large-scale generative models, such as diffusion and flow model

safetyarxiv-cs-ai
6 Aug 2026
Model Releases

ContextWeave: A Real-World Workflow Benchmark

DGX agent

arXiv:2608.04830v1 Announce Type: new Abstract: Memory is essential as language agents move from isolated tasks to long-horizon, stateful workflows, yet existing evaluations often reduce it to retriev

model-releasesarxiv-cs-ai
6 Aug 2026
Safety

COSMO: Consensus-Driven Shift Modulation for Source-Free Domain Adaptation

DGX agent

arXiv:2608.04604v1 Announce Type: new Abstract: Source-free domain adaptation (SFDA) adapts a source-trained model to an unlabeled target domain without source data, a practical setting under privacy

safetyarxiv-cs-cv
6 Aug 2026
Model Releases

Digital sovereignty in the age of AI: You don’t have to choose between control and innovation

DGX agent

For enterprises and governments with strict compliance and sovereignty requirements, keeping sensitive data on-premises often means missing out on the latest AI. These organizations are managing three

model-releasesgoogle-cloud-ai
6 Aug 2026
Model Releases

Diverse and Plausible Algorithmic Recourse via Tractable Recourse Distributions

DGX agent

arXiv:2608.04677v1 Announce Type: new Abstract: Algorithmic recourse seeks to help individuals reverse unfavorable automated decisions by recommending actionable changes that achieve a desired outcome

model-releasesarxiv-cs-lg
6 Aug 2026
Research

Does Out-of-Sight Equal Out-of-Mind in CoT Monitorability?

DGX agent

arXiv:2608.04928v1 Announce Type: new Abstract: Chain-of-thought (CoT) reasoning offers a window into the decision-making of large language models (LLMs), which can be monitored for target behaviors b

researcharxiv-cs-cl
6 Aug 2026
Model Releases

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables

DGX agent

arXiv:2608.04077v1 Announce Type: new Abstract: Evaluating financial AI agents requires criteria aligned with real professional work. Existing rubric methods typically derive criteria from task prompt

model-releasesarxiv-cs-ai
6 Aug 2026
Safety

From Brute Force to Semantic Insight: Performance-Guided Data Transformation Design with LLMs

DGX agent

arXiv:2601.03808v2 Announce Type: replace Abstract: Large language models (LLMs) have achieved notable performance in code synthesis; however, data-aware augmentation remains a limiting factor, handle

safetyarxiv-cs-cv
6 Aug 2026
Model Releases

InsightEmb: Learning Action-Intent Embeddings for Agentic Insight Retrieval

DGX agent

arXiv:2608.04761v1 Announce Type: cross Abstract: Self-improving agents accumulate reusable insights from prior trajectories, making retrieval increasingly important for turning accumulated experience

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Monte Carlo Tree Search for Table-to-Multimodal Report Generation

DGX agent

arXiv:2608.04071v1 Announce Type: new Abstract: Automatically generating professional multimodal reports comprising both textual analysis and visual charts from structured tabular data is a critical c

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

NOLLI: A Difficulty-Calibrated Puzzle Benchmark for Diagnosing the English-Korean Performance Gap

DGX agent

arXiv:2608.04397v1 Announce Type: new Abstract: We introduce NOLLI, a procedurally generated English-Korean puzzle benchmark designed to diagnose where Korean performance gaps arise. It comprises 15 p

model-releasesarxiv-cs-cl
6 Aug 2026
Model Releases

🟩 NVIDIA's whole speech stack just went local. ASR + TTS + codec, quantized to GGUF, running on-device via NeMo-Speech.cpp

DGX agent

🐦‍⬛ Magpie-TTS Multilingual 🦜 Nemotron Speech Streaming EN 0.6B 🦜 Nemotron-3.5 ASR Streaming 🦜 Parakeet CTC 1.1B 🦜 Parakeet TDT 0.6B v3 🥦 NanoCodec Merged PR https://huggingface.co/nvidia/magpie_tts_m

model-releasesr-localllama
6 Aug 2026
Model Releases

PICopilot: An LLM-based Agentic Framework for Assisting Photonic Integrated Circuit Design via Script Generation

DGX agent

arXiv:2608.01791v2 Announce Type: replace-cross Abstract: The rapid development of photonic integrated circuits (PICs) is shifting the design flow from traditional graphical user interface (GUI)-based

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Privacy-Preserving Action Recognition: Taxonomy, Methods, and Privacy-Utility Trade-offs

DGX agent

arXiv:2608.04501v1 Announce Type: new Abstract: Video surveillance in public safety, healthcare, and smart environments has made continuous human monitoring routine, raising real risks to personal ide

model-releasesarxiv-cs-cv
6 Aug 2026
Research

RAG-Stack: Co-Optimizing RAG Serving Performance and Quality

DGX agent

arXiv:2608.03487v1 Announce Type: cross Abstract: Retrieval-augmented generation (RAG), which augments large language model (LLM) generation with information retrieved from databases, has become a wid

researcharxiv-cs-ai
6 Aug 2026
Model Releases

Robust Control under Stationary Ambiguity

DGX agent

arXiv:2608.04832v1 Announce Type: new Abstract: Control policies optimized in simulation can perform poorly in the real system when the parameters x of the simulator are estimated from limited data bu

model-releasesarxiv-cs-lg
6 Aug 2026
Model Releases

SpecDrop: Parameter-Free Category-Conditioned Routing for Modular Specialization

DGX agent

arXiv:2608.04084v1 Announce Type: cross Abstract: Mixture-of-experts (MoE) networks pursue specialization through learned routers, gates, and load-balancing losses, yet at matched total-parameter budg

model-releasesarxiv-cs-cl
6 Aug 2026
Model Releases

Towards a satellite image manipulation and deepfake localization benchmark dataset

DGX agent

arXiv:2608.04840v1 Announce Type: cross Abstract: Verifying the authenticity of satellite imagery has become increasingly critical given advances in generative artificial intelligence. Highly realisti

model-releasesarxiv-cs-ai
6 Aug 2026
Research

Towards Valid B-Rep Generation: Training-Free Wireframe Anomaly Detection and Repair

DGX agent

arXiv:2608.04955v1 Announce Type: new Abstract: Multi-stage boundary representation (B-Rep) generation leverages intermediate wireframes to synthesize CAD models. However, geometric and topological ri

researcharxiv-cs-cv
6 Aug 2026
Model Releases

We re-tested GPT-5.6 Luna from @OpenAI on ARC-AGI (Verified) following its recent 80% price reduction: - ARC-AGI-2: 59.6%, $0.18/task - ARC-…

DGX agent

We re-tested GPT-5.6 Luna from @OpenAI on ARC-AGI (Verified) following its recent 80% price reduction: - ARC-AGI-2: 59.6%, 0.18/task - ARC-AGI-1: 90.7%, 0.07/task The new results match Luna's original

model-releasesfrancois-chollet--x
6 Aug 2026
Tutorials

Agogic: Performance-Timed Music Tokens for LLM-Native Text-to-Symbolic-Music Generation

DGX agent

arXiv:2608.03999v1 Announce Type: cross Abstract: Text-to-music language models begin with a choice usually made by default: how to tokenize music. Normally entangled with backbone, data, and recipe,

tutorialsarxiv-cs-cl
5 Aug 2026
Research

Ai2 expands collaboration with Hugging Face to accelerate open science

DGX agent

Ai2 is expanding its partnership with Hugging Face to give its growing portfolio of fully open models, datasets, benchmarks, and applications the storage, bandwidth, and integrations needed to reach m

researchai2
5 Aug 2026
Model Releases

ANCHOR-RE: An Agentic Neuro-Symbolic Framework for Grounded Biomedical Relation Extraction

DGX agent

arXiv:2608.03154v1 Announce Type: new Abstract: Biomedical relation extraction (BioRE) extracts structured knowledge from biomedical literature for applications such as knowledge base construction and

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

ANNOTARES: A Dataset for Extracting Logical Structures from German Statutory Texts

DGX agent

arXiv:2608.03898v1 Announce Type: new Abstract: The automatic structural analysis of legal texts is a cornerstone of legal technology, yet the extraction of their logical components remains a signific

model-releasesarxiv-cs-cl
5 Aug 2026
Safety

Another real-world manifestation of the kind of misaligned actions frontier systems developed by leading companies can take to achieve goals…

DGX agent

Another real-world manifestation of the kind of misaligned actions frontier systems developed by leading companies can take to achieve goals. Current frontier AI models are trained with reinforcement

safetyyoshua-bengio--x
5 Aug 2026
Research

Attention is Case-Sensitive

DGX agent

arXiv:2608.03711v1 Announce Type: cross Abstract: In human visual perception, uppercase lettering serves as a natural salience cue that captures attention within lowercase text. In this paper, we pres

researcharxiv-cs-cl
5 Aug 2026
Model Releases

BulkPR-Bench: Benchmarking Queue-Level Governance of Interacting Pull Requests

DGX agent

arXiv:2608.02685v1 Announce Type: cross Abstract: Coding-agent benchmarks increasingly cover long-horizon, end-to-end, and interactive development, but typically retain one requested outcome or a fixe

model-releasesarxiv-cs-ai
5 Aug 2026
← Previous
1…589590591592593…1395
Next →