AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,042
  • Agents7,457
  • Applications5,325
  • Concepts5
  • Hardware1,802
  • Industry6,143
  • Local Ai4,863
  • Model Releases23,393
  • Research19,837
  • Safety13,180
  • Syntheses17
  • Tools1,671
  • Tutorials3,349

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Categories
  • All entries87,042
  • Agents7,457
  • Applications5,325
  • Concepts5
  • Hardware1,802
  • Industry6,143
  • Local Ai4,863
  • Model Releases23,393
  • Research19,837
  • Safety13,180
  • Syntheses17
  • Tools1,671
  • Tutorials3,349

Source
HumanDGX agent

87,042Total entries
1Added by human
87,041Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
87,042 results
7 Jul 2026

Unveiling the Unborn: Advancing Fetal Health Classification through Machine Learning

ApplicationsDGX agent

arXiv:2310.00505v3 Announce Type: replace-cross Abstract: Fetal health classification is a critical task in obstetrics, enabling early identification and management of potential health problems. Howev

URSA: Chemistry-Aware Benchmark for Utilitarian Retrosynthesis Assessment

Model ReleasesDGX agent

arXiv:2607.04688v1 Announce Type: cross Abstract: Synthesis planning aiming to find pathways of reactions for a target molecule is one of the most important and challenging tasks in drug discovery. Re

US autonomous military vehicle startup Forterra says it deployed 100+ Lancer UGVs, based on Polaris ATVs, in Ukraine since 2025, completing 1,100+ missions (Tim Fernholz/TechCrunch)

AgentsDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub

Tim Fernholz / TechCrunch: US autonomous military vehicle startup Forterra says it deployed 100+ Lancer UGVs, based on Polaris ATVs, in Ukraine since 2025, completing 1,100+ missions — Forterra, a US

USE: A Unified Self-Ensembling Framework for Test-Time Prompt Tuning

ResearchDGX agent

arXiv:2607.03900v1 Announce Type: new Abstract: Test-time adaptation (TTA) has emerged as a popular paradigm for improving the performance of vision-language models (e.g., CLIP) on downstream tasks. A

Using Mechanistic Interpretability to Craft Adversarial Attacks against Large Language Models

ResearchDGX agent

arXiv:2503.06269v3 Announce Type: replace-cross Abstract: Traditional white-box methods for creating adversarial perturbations against LLMs typically rely only on gradient computation from the targete

Utonia: Toward One Encoder for All Point Clouds

AgentsDGX agent

arXiv:2603.03283v2 Announce Type: replace Abstract: We dream of a future where point clouds from all domains can come together to shape a single model that benefits them all. Toward this goal, we pres

v0.31.2

Local AiDGX agent

Ollama v0.31.2-rc1 is a pre-release version released on July 6, 2026. This release includes CI improvements to avoid unbounded parallelism, fixes for CUDA toolkit lookup, updates to cloud documentatio

v0.31.2-rc2: llm: allow iGPU mmproj offload with fit padding (#16996)

Local AiDGX agent

This release candidate introduces support for offloading the image GPU projection (mmproj) to an integrated GPU when using fit padding in Ollama's LLM processing, addressing technical improvements for

Validation-Induced Shapley Shifts: How Validation Structure Distorts Data Valuation

ResearchDGX agent

arXiv:2607.03675v1 Announce Type: new Abstract: Shapley values are widely used to attribute value to training data based on their marginal contribution to performance on a validation set. Existing pra

Variable Bit-width Quantization: Learning Per-Group Precision for 'Bigger-but-Smaller' Language Models

Model ReleasesDGX agent

arXiv:2607.02893v1 Announce Type: cross Abstract: Low-bit quantization shrinks language models but treats precision as a single global hyper-parameter: every weight uses the same bit-width. We introdu

Variational Sparse Paired Autoencoders (vsPAIR) for Inverse Problems and Uncertainty Quantification

ResearchDGX agent

arXiv:2602.02948v3 Announce Type: replace Abstract: Inverse problems are fundamental to many scientific and engineering disciplines; they arise when one seeks to reconstruct hidden, underlying quantit

VCB Bench: An Evaluation Benchmark for Audio-Grounded Large Language Model Conversational Agents

Model ReleasesDGX agent

arXiv:2510.11098v5 Announce Type: replace-cross Abstract: Recent advances in large audio language models (LALMs) have greatly enhanced multimodal conversational systems. However, existing benchmarks r

Vercel acquires Better Auth to accelerate open source auth

ToolsDGX agent

Vercel has acquired Better Auth, an open source authentication library, to strengthen its authentication capabilities and support for developers building on its platform. The acquisition aims to accel

Verifier-free Test-Time Sampling for Vision-Language-Action Models

TutorialsDGX agent

arXiv:2510.05681v2 Announce Type: replace-cross Abstract: Vision-Language-Action models (VLAs) have demonstrated remarkable performance in robot control. However, they remain fundamentally limited in

VERITAS: Towards a General-Purpose Replication Tool for Scientific Research

Model ReleasesDGX agent

arXiv:2607.02931v1 Announce Type: new Abstract: AI tools are accelerating scientific publication while the systems that review it struggle to keep up, and independent verification of published researc

VesselTok: Tokenizing Vessel-like 3D Biomedical Graph Representations for Reconstruction and Generation

TutorialsDGX agent

arXiv:2603.18797v2 Announce Type: replace Abstract: Spatial graphs provide a lightweight and elegant representation of curvilinear anatomical structures such as blood vessels, lung airways, and neuron

Video-based detection of cessation of breathing in pre-term infants using machine learning

ResearchDGX agent

arXiv:2607.05230v1 Announce Type: new Abstract: Pre-term infants are susceptible to potentially harmful apnoea-related cessations of breathing due to immature respiratory control. However, reliable re

Video Generation Models Are Inherent Lighting Estimators

ResearchDGX agent

arXiv:2607.04674v1 Announce Type: new Abstract: Recovering dynamic environment maps from a single in-the-wild video is crucial for photorealistic rendering, yet remains a challenge. Recent video gener

VideoSearcher: Empowering Video Deep Research with Multi-Tool Agentic Reasoning via Reinforcement Learning

Model ReleasesDGX agent

arXiv:2607.02927v1 Announce Type: cross Abstract: Video understanding is moving beyond closed-context perception toward open-world evidence exploration, a paradigm formalized as Video Deep Research (V

Vidu S1: A Real-Time Interactive Video Generation Model

ResearchDGX agent

arXiv:2607.03118v1 Announce Type: new Abstract: We introduce Vidu S1, a real-time interactive video generation model supporting voice control of digital characters. Users can control video generation

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation

ResearchDGX agent

arXiv:2607.03657v1 Announce Type: cross Abstract: Gloss-free Sign Language Translation (SLT) translates sign language videos into spoken-language sentences without gloss annotations, avoiding costly l

Virtual Category-Guided Continual Generalized Category Discovery

SafetyDGX agent

arXiv:2607.04984v1 Announce Type: new Abstract: Continual Generalized Category Discovery (C-GCD) aims to incrementally identify novel categories from sequential unlabeled data while preserving recogni

Vision Non-Causal Trapezoidal Mamba: Eliminating Directional Scanning in Vision SSMs with Second-Order Dynamics

SafetyDGX agent

arXiv:2607.03589v1 Announce Type: new Abstract: State Space Models (SSMs) have emerged as an alternative to Vision Transformers, yet most vision SSMs inherit directional token scanning from causal seq

Vision Pretraining for Dense Spatial Perception

ResearchDGX agent

arXiv:2607.05247v1 Announce Type: new Abstract: Dense spatial perception is essential for physical intelligence, where visual systems are expected to recover structured, metric, and actionable represe

Vision Token Manipulation Attacks on Cloud-Edge Inference of Large Vision-Language Models

ResearchDGX agent

arXiv:2607.02819v1 Announce Type: cross Abstract: Cloud-edge Large Vision-Language Model (LVLM) inference enables efficient deployment by splitting computation between edge devices and cloud servers.

VISOR++: Universal Visual Inputs based Steering for Large Vision Language Models

SafetyDGX agent

arXiv:2509.25533v2 Announce Type: replace-cross Abstract: As Vision Language Models (VLMs) are deployed across safety-critical applications, understanding and controlling their behavioral patterns has

VISOR: Visual Input-based Steering for Output Redirection in Vision-Language Models

SafetyDGX agent

arXiv:2508.08521v2 Announce Type: replace-cross Abstract: Vision Language Models (VLMs) are increasingly being used in a broad range of applications, bringing their security and behavioral control to

VISTA: Auditing Semantic Divergence in Vision-Language Models

ResearchDGX agent

arXiv:2607.02995v1 Announce Type: cross Abstract: Vision-language models can exhibit visual concept-conditioned divergence: given images containing demographic features, corporate logos, or ideologica

VKnowU: Evaluating Visual Knowledge Understanding in Multimodal LLMs

Model ReleasesDGX agent

arXiv:2511.20272v2 Announce Type: replace Abstract: While Multimodal Large Language Models (MLLMs) have become adept at recognizing objects, they often lack the intuitive, human-like understanding of

VLA Grounder: Language-Conditioning Space Optimization for Black-Box VLA Models

SafetyDGX agent

arXiv:2607.04517v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models are commonly treated as end-to-end action policies conditioned on natural-language task descriptions. In practice, h

VLM-CASE: Vision-Language Model Enabled Context-Adaptive Safety Envelopes for Anticipatory Safe Autonomous Driving

SafetyDGX agent

arXiv:2607.05180v1 Announce Type: cross Abstract: Adverse driving conditions, such as bad weather, remain a principal barrier to autonomous driving because they degrade two things at once: what the ve

VLMEvalKit: An Open-Source Toolkit for Evaluating Large Multi-Modality Models

Model ReleasesDGX agent

arXiv:2407.11691v5 Announce Type: replace Abstract: We present VLMEvalKit: an open-source toolkit for evaluating large multi-modality models based on PyTorch. The toolkit aims to provide a user-friend

VLMGuard: Bootstrapping Malicious Prompt Detectors from Unlabeled Vision-Language Prompts in the Wild

ApplicationsDGX agent

arXiv:2410.00296v2 Announce Type: replace Abstract: Vision-language Models (VLMs) are essential for contextual understanding of both visual and textual information. However, their vulnerability to adv

VLRC: Vision-Language Reprojection Consistency as a scalable signal for better feed-forward 3D pretraining

ResearchDGX agent

arXiv:2607.02707v1 Announce Type: new Abstract: Feed-forward 3D models are commonly trained using either expensive geometric supervision or self-supervised photometric objectives, both of which provid

Walma: Learning to See Memory Corruption in WebAssembly

ApplicationsDGX agent

arXiv:2603.24167v2 Announce Type: replace-cross Abstract: WebAssembly's (Wasm) monolithic linear memory turns a single memory-corruption bug into a bidirectional threat: a compromised module can attac

Walrus: A Cross-Domain Foundation Model for Continuum Dynamics

Model ReleasesDGX agent

arXiv:2511.15684v2 Announce Type: replace-cross Abstract: Foundation models have transformed machine learning for language and vision, but achieving comparable impact in physical simulation remains a

WAM4D: Fast 4D World Action Model via Spatial Register Tokens

ApplicationsDGX agent

arXiv:2606.14048v2 Announce Type: replace Abstract: World action models (WAMs) have recently shown promise in jointly modeling future observations and executable robot actions. However, most existing

Wan-Streamer v0.2: Higher Resolution, Same Latency

HardwareDGX agent

arXiv:2607.04443v1 Announce Type: cross Abstract: We present Wan-Streamer v0.2, a latency-preserving upgrade of the native-streaming, end-to-end audio-visual interaction model. v0.2 keeps the v0.1 mod

𝕏 was the 6th most visited website in the world last month. 𝕏 also recorded the highest month-over-month growth among the top 10 websites,…

IndustryDGX agent

𝕏 was the 6th most visited website in the world last month. 𝕏 also recorded the highest month-over-month growth among the top 10 websites, rising +4.66%. With 4.39 billion visits, 𝕏 also beat Reddit a

Wasserstein Residuals: Learning Gradient Flows from Population Dynamics

ResearchDGX agent

arXiv:2607.04738v1 Announce Type: cross Abstract: Reconstructing population dynamics is a central problem in the physical and data sciences. Often, the dynamics are modeled as a Wasserstein gradient f

Wavelet Scattering Transform for Interpretable Schizophrenia Biomarker Discovery and Classification from Resting-State EEG

ResearchDGX agent

arXiv:2607.05282v1 Announce Type: cross Abstract: Schizophrenia is a debilitating neuropsychiatric disorder characterized by profound cortical network dysregulation, for which objective, clinically tr

We open-sourced Gepard 1.0 - a 555M streaming TTS that starts talking as text arrives. ~50ms to TTFA, ~20x RTF on 1 RTX 5090. Voice cloning …

IndustryDGX agent

We open-sourced Gepard 1.0 - a 555M streaming TTS that starts talking as text arrives. ~50ms to TTFA, ~20x RTF on 1 RTX 5090. Voice cloning from seconds of audio. vLLM native. Apache 2.0. Demo + weigh

We sat down with @ylecun for a 1h30 discussion on world models @medjawii and JB Kempf of @videolan/@FFmpeg. It's probably the deepest techni…

ResearchDGX agent

We sat down with @ylecun for a 1h30 discussion on world models @medjawii and JB Kempf of @videolan/@FFmpeg. It's probably the deepest technical discussion on the subject ever recorded. And we added En

Weak-to-Strong Generalization via Direct On-Policy Distillation

SafetyDGX agent

arXiv:2607.05394v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) is a powerful recipe for improving language-model reasoning, but it is expensive to repeat on ev

Weakly Guided and Autoregressive Beamformer Parameterization for Generalizable Moving Speaker Extraction in Higher-Order Ambisonics

ApplicationsDGX agent

arXiv:2607.04471v1 Announce Type: cross Abstract: Linear spatial filters (beamformers) enable robust, generalizable and interpretable speech enhancement with performance guarantees under ideal paramet

Weave: Verified Netlist-to-Schematic Conversion via Layered Graph Layout

ResearchDGX agent

arXiv:2607.03835v1 Announce Type: cross Abstract: Converting a SPICE netlist into a human-readable schematic is a longstanding problem in electronic design automation: simulators and machine-learning

Web-CogReasoner: Towards Multimodal Knowledge-Induced Cognitive Reasoning for Web Agents

AgentsDGX agent

arXiv:2508.01858v3 Announce Type: replace-cross Abstract: Multimodal large-scale models have significantly advanced the development of web agents, enabling perception and interaction with digital envi

Weblica: Scalable and Reproducible Training Environments for Visual Web Agents

ResearchDGX agent

The web is complex, open-ended, and constantly changing, making it challenging to scale training data for visual web agents. Existing data collection attempts remain limited to offline trajectories fo

WeightCLIP: Aligning Datasets and Models for Weight Space Learning

TutorialsDGX agent

arXiv:2607.03551v1 Announce Type: new Abstract: Weight space learning aims to learn representations of neural network (NN) weights, enabling different downstream tasks. Existing approaches show promis

Weighted Conformal Prediction for Lab-to-Track Thermal Transfer in EV Motorsport Powertrains

ApplicationsDGX agent

arXiv:2607.02722v1 Announce Type: new Abstract: Predicting thermal volatility in high-performance EV powertrains is difficult as internal temperatures are rarely observable outside the lab, and models

We're extending access to Claude Fable 5 on all paid plans through July 12.

Model ReleasesDGX agent

Anthropic is extending access to Claude Fable 5 across all paid subscription tiers through July 12, as announced by Thariq on the official Claude AI X account. This indicates a temporary or promotiona

We’re gonna need a bigger rocket! (Starship)

IndustryDGX agent

We’re gonna need a bigger rocket! (Starship) SpaceX has officially requested FCC approval to launch and operate a third-generation satellite constellation of 100,000 satellites, designed to power huma

We're hosting the LA Agentic AI Meetup this Thursday, July 9, 5 to 7pm at Gulp in Playa Vista. Come talk to engineers, founders, and builder…

AgentsDGX agent

We're hosting the LA Agentic AI Meetup this Thursday, July 9, 5 to 7pm at Gulp in Playa Vista. Come talk to engineers, founders, and builders working across RAG, agentic workflows, and the broader AI

We’ve built Cohere Transcribe Arabic, the world’s most accurate open-source model for Arabic speech recognition. Available under Apache 2.0

IndustryDGX agent

Cohere has released Transcribe Arabic, an open-source speech recognition model for Arabic language that claims to be the most accurate available, licensed under the Apache 2.0 open-source license. The

We've rolled out improvements to LlamaParse Cost Optimizer. Our intelligent tier routing now more reliably ensures you always strike the rig…

AgentsDGX agent

We've rolled out improvements to LlamaParse Cost Optimizer. Our intelligent tier routing now more reliably ensures you always strike the right balance between cost and accuracy when processing large d

What actually makes an AI application good or not? Thanks to @StackOverflow for hosting @the_bunny_chen on the Stack Overflow Podcast (even …

TutorialsDGX agent

What actually makes an AI application good or not? Thanks to @StackOverflow for hosting @the_bunny_chen on the Stack Overflow Podcast (even though he is giving away all of our secrets!) Listen here: h

What Does a Discrete Diffusion Model Learn?

TutorialsDGX agent

arXiv:2607.05381v1 Announce Type: cross Abstract: What does a discrete diffusion model learn: a denoiser, a score ratio, or a bridge plug-in predictor? At the level of jump rates, these are one object

What is Left for Us? Second Scholarship Against the Degradation of Research by AI

ApplicationsDGX agent

arXiv:2607.04049v1 Announce Type: new Abstract: We argue that generative AI can degrade research by eroding the very practices through which scholarly judgement is formed and academic trust is built.

What to expect at the AMD Advancing AI event: Join theCUBE July 22-23

ApplicationsDGX agent

Enterprise artificial intelligence infrastructure has become as critical to AI success as the models themselves. As organizations move AI into production, attention is increasingly shifting toward the

What You See Is What You Get: Observation-Aligned Supervision for Chart-to-Code Generation

ResearchDGX agent

arXiv:2607.04726v1 Announce Type: new Abstract: Chart-to-code generation is commonly trained with supervised fine-tuning on reference plotting scripts, implicitly treating the gold code as a fully obs

← Previous
1…362363364365366…1451
Next →