AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,115
  • Agents7,313
  • Applications5,228
  • Concepts5
  • Hardware1,762
  • Industry6,105
  • Local Ai4,756
  • Model Releases22,759
  • Research19,333
  • Safety12,889
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Categories
  • All entries85,115
  • Agents7,313
  • Applications5,228
  • Concepts5
  • Hardware1,762
  • Industry6,105
  • Local Ai4,756
  • Model Releases22,759
  • Research19,333
  • Safety12,889
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

85,115Total entries
1Added by human
85,114Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
85,115 results
24 Jul 2026

Uncertainty-Aware Trust Estimation for Multi-LLM Systems via Structured Expert Judgement

ResearchDGX agent

arXiv:2607.20529v1 Announce Type: cross Abstract: Large Language Model (LLM) ensembles are increasingly used to improve reliability by combining predictions from multiple LLMs. However, existing aggre

UnDA: Unpaired Domain Alignment for Cross-Modal Knowledge Transfer in Medical Imaging

SafetyDGX agent

arXiv:2607.21546v1 Announce Type: new Abstract: Multimodal based approaches often outperform single modality approaches in downstream tasks as the different modalities provide complementary informatio

Understanding Critical Thinking in Generative Artificial Intelligence Use: Development, Validation, and Correlates of the Critical Thinking in AI Use Scale

SafetyDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub

arXiv:2512.12413v2 Announce Type: replace Abstract: Generative AI tools are increasingly embedded in everyday work and learning, yet their fluency, opacity, and propensity to hallucinate mean that use

Unified Video Dense Prediction from Disjoint Data

ResearchDGX agent

arXiv:2607.21592v1 Announce Type: new Abstract: Scene understanding requires simultaneous prediction about geometry, appearance, and semantics. However, existing task-specific annotations are fragment

Unlearning Under Imbalance: Benchmarking Fairness in Multimodal LLM Unlearning

Model ReleasesDGX agent

arXiv:2607.21300v1 Announce Type: cross Abstract: Machine unlearning has emerged as a tool for removing personal data from trained models to comply with recent AI regulations. To evaluate unlearning e

Unsupervised Consensus-Based Anomaly Detection for Spatiotemporal Malaria Incidence in Ghana

ResearchDGX agent

arXiv:2607.21559v1 Announce Type: new Abstract: A consensus anomaly detection framework was applied to monthly malaria surveillance data from Ghana (2014-2023) to identify atypical transmission patter

Unsupervised Metal Artifact Reduction in Dental CBCT using Fine-tuned Cycle-Consistent Adversarial Networks

ResearchDGX agent

arXiv:2607.20977v1 Announce Type: new Abstract: Metal artifacts generated by dental implants significantly degrade cone-beam computed tomography (CBCT) volumes, obscuring critical anatomical structure

URF: A Unified Robot Control-Policy Framework for Stable Contact Aware Manipulation

SafetyDGX agent

arXiv:2607.20912v1 Announce Type: new Abstract: Learning-based manipulation policies usually predict robot actions from sensory observations and leave their execution to a separate low-level controlle

V-DEAL: Diagnosing Video Safety De-Calibration as an Understanding-Refusal Coupling Failure

SafetyDGX agent

arXiv:2607.21151v1 Announce Type: new Abstract: As Video Large Language Models are increasingly deployed in real-world applications, ensuring their safety alignment has become critical. Counterintuiti

Verifier-First Evaluation of Agentic LLMs for Infrastructure-as-Code Generation

Model ReleasesDGX agent

arXiv:2607.20478v1 Announce Type: cross Abstract: Infrastructure-as-Code (IaC) generation from natural language requires satisfying provider schemas, dependency planning, and organizational policy con

VeriSimpl: Robust Optimization Modeling from Natural Language using Simplification-based Verification

ResearchDGX agent

arXiv:2607.20474v1 Announce Type: new Abstract: Natural language interfaces can greatly benefit the accessibility and usability of optimization modeling, and recent advances in large language models (

VibeVoice-ASR-BitNet Technical Report

Local AiDGX agent

arXiv:2607.21075v1 Announce Type: cross Abstract: We present VibeVoice-ASR-BitNet, a compressed variant of VibeVoice-ASR optimized for real-time inference on edge CPUs. We apply heterogeneous quantiza

Vision-Language-Policy Model for Dynamic Robot Task Planning

SafetyDGX agent

arXiv:2512.19178v2 Announce Type: replace-cross Abstract: Bridging the gap between natural language commands and autonomous execution in unstructured environments remains an open challenge for robotic

ViSTR-Bench: Can MLLMs Reason from Continuous Visual Cues in Dynamic Scenes?

Model ReleasesDGX agent

arXiv:2607.20868v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have achieved remarkable success across diverse expert-level tasks, but they still struggle with fundamental ab

Visual Contrastive Self-Distillation

Model ReleasesDGX agent

arXiv:2607.21556v1 Announce Type: cross Abstract: On-policy self-distillation (OPSD) is promising as it removes the external teacher required by on-policy distillation (OPD), yet it still needs asymme

VoLN: Vision-Only Long-Horizon Navigation---Paradigm, Benchmark, and Method

Model ReleasesDGX agent

arXiv:2607.21400v1 Announce Type: cross Abstract: Vision-and-Language Navigation (VLN) enables embodied agents to follow natural-language instructions. However, route-level instructions commonly encod

VPWEM: Non-Markovian Visuomotor Policy with Working and Episodic Memory

Model ReleasesDGX agent

arXiv:2603.04910v2 Announce Type: replace-cross Abstract: Imitation learning from human demonstrations has achieved significant success in robotic control, yet most visuomotor policies still condition

Was great to talk about this timely and alarming news, thanks for having me on!

SafetyDGX agent

Was great to talk about this timely and alarming news, thanks for having me on! . @NPCollapse, executive director at ControlAI, joins 'On Balance' to discuss the dangers of artificial intelligence aft

WAT3R: Feedforward Underwater 3D Reconstruction

ResearchDGX agent

arXiv:2607.21023v1 Announce Type: new Abstract: Reliable feedforward underwater 3D reconstruction remains challenging due to severe light attenuation and backscattering, which degrade visual quality a

WaveformQA: Benchmarking LLM Temporal Reasoning on Digital Waveforms

Model ReleasesDGX agent

arXiv:2607.20638v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated strong capabilities in code generation and reasoning, yet their ability to perform temporal reasoning ove

We are pleased to announce that @GaryMarcus, Professor Emeritus at New York University, will speak at the 19th Annual AGI Conference (July 2…

SafetyDGX agent

We are pleased to announce that @GaryMarcus, Professor Emeritus at New York University, will speak at the 19th Annual AGI Conference (July 27–30, 2026). Professor Marcus is a leading voice in artifici

We benchmarked Opus 5 comprehensively on document understanding through ParseBench. It is roughly on par with Opus 4.8 - it does a few point…

Model ReleasesDGX agent

We benchmarked Opus 5 comprehensively on document understanding through ParseBench. It is roughly on par with Opus 4.8 - it does a few points worse on dense tables, but does slightly better on parsing

WebCoach: Self-Evolving Web Agents with Cross-Session Memory Guidance

Model ReleasesDGX agent

arXiv:2511.12997v2 Announce Type: replace Abstract: Multimodal LLM-powered agents have recently demonstrated impressive capabilities in web navigation, enabling agents to complete complex browsing tas

Webly Supervised Multi-Label Recognition: Evaluation Benchmark and Dual-Branch Multi-Label Contrastive Learning

Model ReleasesDGX agent

arXiv:2607.20874v1 Announce Type: new Abstract: Training deep learning models with freely available web images can reduce their dependence on costly manual annotations. Although webly supervised learn

Weight-norm Criticality: A Mechanism for Loss Spikes Induced by the Normalization and Weight Decay

Model ReleasesDGX agent

arXiv:2607.21005v1 Announce Type: new Abstract: Most explanations of training instability focus on learning-rate criticality, typically characterized by the Edge of Stability, beyond which optimizatio

Welcome @jensenhuang 💚 The future of AI leadership needs both frontier closed models and frontier open models.

SafetyDGX agent

Welcome @jensenhuang 💚 The future of AI leadership needs both frontier closed models and frontier open models. For my first post, I’m sharing a letter @NVIDIA signed on why open models matter. AI will

What I learned using Ollama on a real Paperless archive: model choice was not the main problem

Local AiDGX agent

I maintain Tagvico, an open-source companion for Paperless-ngx. I added Ollama because document text is exactly the kind of data many people do not want to send to a hosted model. The surprising failu

What if AI had access to classified files it was never allowed to quote but can make image?

IndustryDGX agent

What if an AI had seen fragments of classified material it could never describe directly? No files. No report names. No official explanations. Just images. That was the concept behind this series. I a

What is Good? Extracting and Testing Implicit Theories of Literary Quality from LLM Reasoning Traces

Model ReleasesDGX agent

arXiv:2607.20425v1 Announce Type: new Abstract: What makes writing 'good' remains a persistent question in literary studies and computational linguistics. We present a two-study investigation of how r

What Matters for Simulation to Online Reinforcement Learning on Real Robots

ApplicationsDGX agent

arXiv:2602.20220v2 Announce Type: replace-cross Abstract: We investigate what specific design choices enable successful online reinforcement learning (RL) on physical robots. Across 100 real-world tra

What, Where, and How: Disentangling the Roles of Task, Language, and Model in Code Model Representations

Model ReleasesDGX agent

arXiv:2607.21491v1 Announce Type: new Abstract: Do independently trained language models come to represent the same thing in the same way? We answer for code, extending a recently introduced concept-c

What's the last model trained on human-data only?

Local AiDGX agent

From my understanding, most current LLMs are trained on trillions and trillions of tokens of mostly AI-generated data. Are there any recent models that are trained purely (or as close as possible) on

When Are Reasoning-Based Guardrails Not Efficient? ResponseGuard: A Fast Vision-Language Guard for Real-Time Moderation

Model ReleasesDGX agent

arXiv:2607.21401v1 Announce Type: cross Abstract: A vision-language AI assistant returns its answer as a stream of generated tokens. Therefore, a safety guard that watches that answer has to keep up w

When Does Recurrence Become an Algorithm? Convergence Selection in Weight-Tied Looped Transformers

Model ReleasesDGX agent

arXiv:2607.20594v1 Announce Type: cross Abstract: When does a weight-tied looped transformer -- one block applied T times -- implement an actual algorithm? We answer with four findings from controlled

When RLVR Shrinks the Reasoning Boundary: Diagnosing Pass@k Inversion

Model ReleasesDGX agent

arXiv:2607.20543v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) can improve one-sample accuracy while making a model worse under repeated sampling. We study thi

When Trivia Is Not Trivial: Everyday Knowledge Failures in Multilingual LLMs

Model ReleasesDGX agent

arXiv:2607.21445v1 Announce Type: new Abstract: Quiz rooms, trivia nights, and quiz shows challenge human knowledge across a wide range of topics, from canonical facts to everyday culture. In this pap

Where Animacy Lives in Large Language Models: Tracing the Circuits of the Animacy Concept

ResearchDGX agent

arXiv:2607.20995v1 Announce Type: new Abstract: Distinguishing animate from inanimate concepts in written language requires more than shallow text processing, as it involves recognizing complex select

WhereEdit: Mask-aware Local Latent Editing for One-Step Image Editing

Model ReleasesDGX agent

arXiv:2607.20883v1 Announce Type: new Abstract: Recent one-step text-to-image (T2I) models enable efficient image synthesis and provide new opportunities for real-time image editing. However, existing

which model to use on local 24gb mac mini M4 pro

Local AiDGX agent

So, i have been building some apps that should run on the local every user system, tried gemma4 although its fast and great at reasoning its not as good in instructions following and tool calling. tri

Who has set up ChatGPT Finance?

ApplicationsDGX agent

It’s been out for Plus members for a little while now. Has anyone here connected it to their accounts? I myself haven’t done it, even though it’s only read access I’m seriously hesitant to hand over t

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning

Model ReleasesDGX agent

arXiv:2607.09328v2 Announce Type: replace-cross Abstract: Answering complex questions over long documents frequently requires integrating evidence that the source itself disperses naturally across dis

Will there be Flux 3 Klein?

SafetyDGX agent

https://bfl.ai/blog/flux-3 “Over the next few weeks and months, we will make the following capabilities available, each after an early access phase for ensuring smooth rollout, collecting feedback and

Windowed-MTP: Removing the Full-Context Draft-KV Tax at Million-Token Context

Model ReleasesDGX agent

arXiv:2607.21535v1 Announce Type: cross Abstract: Speculative decoding accelerates autoregressive generation by having a cheap draft propose tokens that a target verifies in parallel. Frontier models

Wireless TokenCom: RL-Based Tokenizer Agreement for Multi-User Wireless Token Communications

SafetyDGX agent

arXiv:2602.12338v2 Announce Type: replace Abstract: Token Communications (TokenCom) has recently emerged as an effective new paradigm, where tokens are the unified units of multimodal communications a

Wisdom of LLM Crowds: Aggregation and Contamination in Language Model Ensembles

Local AiDGX agent

arXiv:2607.18269v2 Announce Type: replace Abstract: The wisdom of crowds -- the finding that aggregating judgments across individuals often outperforms the best individual -- has been extensively stud

Word meaning co-determines vowel-inherent spectral change. A corpus-based investigation of conversational Mandarin

ApplicationsDGX agent

arXiv:2607.21391v1 Announce Type: new Abstract: This study investigates vowel-inherent spectral change (VISC) in spontaneous conversational Mandarin. Using the generalized additive model and word embe

Workflow-Localized Mechanism Learning: Attribution-Guided Repair and Knowledge Reuse for Structured Agent Skills

Model ReleasesDGX agent

arXiv:2607.20999v1 Announce Type: new Abstract: Agent Skills package reusable procedural knowledge as external artifacts for frozen language-model agents, yet existing optimizers do not jointly resolv

Workload-Aware Caching for Multi-Agent Systems

SafetyDGX agent

arXiv:2607.20495v1 Announce Type: new Abstract: Multi-agent systems decompose complex tasks into directed acyclic graphs (DAGs) of specialized agent executions, creating natural opportunities for cach

Writhe-Based Polymer Link Classification Using Machine Learning

ResearchDGX agent

arXiv:2607.20657v1 Announce Type: cross Abstract: Unique and rapid classification of knots and links is an open mathematical problem that is relevant to a range of (bio)physical systems, including pol

X^3-OPD: Distilling Reasoning into Large Audio-Language Models via On-Policy Alignment

SafetyDGX agent

arXiv:2607.21550v1 Announce Type: new Abstract: While large audio-language models have achieved remarkable progress in auditory perception, they still lag behind text-based large language models in de

Zagreus-0.4B-por a small open source language model for Portuguese

Model ReleasesDGX agent

mii-llm, an open source AI lab, released Zagreus-0.4B-por, a compact bilingual Portuguese–English language model pretrained entirely from scratch. The model has approximately 400 million parameters an

Zero-Flow Two-Sample Tests

TutorialsDGX agent

arXiv:2607.21542v1 Announce Type: new Abstract: We propose a new approach to two-sample testing for deciding whether two sets of samples are drawn from the same distribution. The test is built on a st

ZONDA: Zero-shot Object Navigation with Dynamic Avoidance in Multi-floor Environments

Model ReleasesDGX agent

arXiv:2607.21025v1 Announce Type: new Abstract: In Object Goal Navigation task, existing methods are typically restricted to static and single-floor environments, ignoring cross-floor topologies and d

23 Jul 2026

4DGS360: 360{eg} Gaussian Reconstruction of Dynamic Objects from a Single Video

Model ReleasesDGX agent

arXiv:2603.21618v2 Announce Type: replace Abstract: We introduce 4DGS360, a diffusion-free framework for 360^{irc} dynamic object reconstruction from casual monocular video. Existing methods often fai

A Bayesian Framework for Built-in Input Dimension Reduction for Gaussian Process Modeling

ResearchDGX agent

arXiv:2607.19498v1 Announce Type: cross Abstract: Gaussian process (GP) modeling is widely used in computational science and engineering. However, fitting a GP to high-dimensional inputs remains chall

A caveman qwen3.6 27B

Local AiDGX agent

Just saw this on huggingface: https://huggingface.co/ProCreations/grug-27b The benchmarks claim that it's quite a bit better than qwen3.6 27B original and that they reduced the amount of necessary tok

A Confidence Interval for the ell_2 Expected Calibration Error

ResearchDGX agent

arXiv:2408.08998v4 Announce Type: replace-cross Abstract: Recent advances in machine learning have significantly improved prediction accuracy in various applications. However, ensuring the calibration

A convergence result of a continuous model of deep learning via a L{}ojasiewicz--Simon inequality

Model ReleasesDGX agent

arXiv:2311.15365v3 Announce Type: replace Abstract: We study an idealized training process for deep neural networks in a continuous-depth, mean-field model in which each layer is parameterized by a pr

A Deep Learning Framework for Predicting Solar EUV Irradiance During Significant Flares

ResearchDGX agent

arXiv:2607.19597v1 Announce Type: cross Abstract: We present FlareEUV, a multimodal deep learning framework for predicting daily extreme ultraviolet (EUV) irradiance at 6.5 nm over three consecutive d

A Framework of User Experience Principles for Human-AI Agent Interaction in the Workplace

AgentsDGX agent

arXiv:2607.19941v1 Announce Type: cross Abstract: As AI agents become integral to business workflows, establishing guiding user experience (UX) principles is crucial for ensuring user trust and succes

← Previous
1…223224225226227…1419
Next →