AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
Human
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
84,459 results
6 Aug 2026

DeepInvert: Semi-Supervised Embedding Inversion Against Obfuscated Language Models

ResearchDGX agent

arXiv:2608.04477v1 Announce Type: cross Abstract: Cloud-based language model services routinely process prompts containing sensitive information. Obfuscation-based defenses---including ObfusLM, Sentin

DeepSeek says it plans to implement substantial price increases across its services; V4 Flash currently costs 0.14/1M input and 0.28/1M output tokens (Bloomberg)

Model ReleasesDGX agent

Bloomberg: DeepSeek says it plans to implement substantial price increases across its services; V4 Flash currently costs 0.14/1M input and 0.28/1M output tokens — DeepSeek plans to implement a signifi

DeepSeek-V4 Flash 0731 vs GPT-5.6 Luna on DeepSWE: Cost and Coding

Model Releases
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

DeepSeek‑V4 Flash 0731 is the cheapest model on the DeepSWE board, costing about 0.10 per rollout versus GPT‑5.6 Luna’s 0.61, yet it scores a pass@1 of 53.3% compared to Luna’s 67.2%. A cascade strate

Deliberate Before You Fly: Vision-Guided Spatial Deliberation for UAV See-and-Reach Navigation

AgentsDGX agent

arXiv:2608.04825v1 Announce Type: new Abstract: UAV see-and-reach navigation requires an aerial agent to approach a language-specified target visible in its initial view and stop reliably near it. Exi

Delinea targets AI agent risks with real-time authorization

AgentsDGX agent

As AI agents gain access to sensitive enterprise systems, AI agent security is becoming increasingly difficult to manage with traditional identity and privilege controls. The shift is driving demand f

Deltoris: Enabling Real-time VLA Inference in Embodied AI via Bit-level Sparsity and Speculative Inference

ResearchDGX agent

arXiv:2608.04428v1 Announce Type: cross Abstract: Vision-language-action (VLA) models have emerged as a key component in embodied AI. Among existing approaches, diffusion-based VLA models achieve supe

DelusionEval: Measuring Delusion-Linked Behaviors in AI Chatbots

Model ReleasesDGX agent

arXiv:2608.05004v1 Announce Type: new Abstract: Mental health professionals have raised concerns about risks of psychological harm from interaction with large language models (LLMs), including 'delusi

Dense Metric Depth Completion from Sparse Direct Time-of-Flight Sensors

TutorialsDGX agent

arXiv:2608.04737v1 Announce Type: new Abstract: Direct Time-of-Flight (dToF) sensors provide highly accurate metric depth and are more robust than indirect ToF systems in challenging real-world condit

Design and Flight of an Ion-propelled Micro Hovercraft Leveraging Ground Proximity Effects

AgentsDGX agent

arXiv:2608.04343v1 Announce Type: new Abstract: Electroaerodynamic propulsion is compelling for use in micro air vehicles due to its silent and solid-state nature, but its limited efficiency has thus

Design Choices That Matter: A Functional ANOVA Analysis for Remote Sensing Multi-Label Classification

ResearchDGX agent

arXiv:2608.04702v1 Announce Type: cross Abstract: Benchmarking deep learning (DL) models for multi-label classification (MLC) of remote sensing images (RSI) typically yields rankings that do not gener

Diagnosing Tool-Selection Reasoning in LLM Agents with Canary Tools

Model ReleasesDGX agent

arXiv:2608.04719v1 Announce Type: new Abstract: Agent evaluations tell us that a model picked the wrong tool, but rarely why. We introduce canary tools: diagnostic probe tools planted in an agent's Mo

Differential 6-DOF Pose Estimation with Provable First-Order Immunity to Camera Calibration Errors

Model ReleasesDGX agent

arXiv:2608.04673v1 Announce Type: new Abstract: Accurate six-degree-of-freedom (6-DOF) motion estimation is essential for robotic manipulation, autonomous systems, and structural displacement monitori

Differentiating Through Dual Prices: End-to-End Policy Learning Under Capacity Constraints

SafetyDGX agent

arXiv:2608.04669v1 Announce Type: new Abstract: Many social services assign scarce resources, such as housing assistance or hospital interventions, to people who arrive one at a time: each arrival mus

Digital sovereignty in the age of AI: You don’t have to choose between control and innovation

Model ReleasesDGX agent

For enterprises and governments with strict compliance and sovereignty requirements, keeping sensitive data on-premises often means missing out on the latest AI. These organizations are managing three

Discretization and Statistical Consistency of Functional Flow Matching

ResearchDGX agent

arXiv:2608.04531v1 Announce Type: new Abstract: Functional flow matching is posed on distributions of functions but implemented from finitely many coefficients or point values. Under scattered or adap

DisMix: Order-Aware Mixup for Medical Imaging via Disentangling Ordinal and Non-Ordinal Features

ResearchDGX agent

arXiv:2608.04652v1 Announce Type: cross Abstract: Image mixup is a widely adopted data augmentation strategy, yet it is ill-suited for ordinal classification tasks such as medical disease grading, whe

Distributional Active Inference

ResearchDGX agent

arXiv:2601.20985v2 Announce Type: replace Abstract: Optimal control of complex environments with robotic systems faces two complementary and intertwined challenges: efficient organization of sensory s

DIVE: Dynamic Iterative Visual Evidence Construction for Efficient Vision-Language Models

ResearchDGX agent

arXiv:2608.04496v1 Announce Type: new Abstract: Visual inputs in vision-language models (VLMs) are often encoded into substantially longer token sequences than text, making visual tokens a major bottl

Diverse and Plausible Algorithmic Recourse via Tractable Recourse Distributions

Model ReleasesDGX agent

arXiv:2608.04677v1 Announce Type: new Abstract: Algorithmic recourse seeks to help individuals reverse unfavorable automated decisions by recommending actionable changes that achieve a desired outcome

Do Language Models Know Their Slang? Queer Slang Understanding in User-Generated Content

ResearchDGX agent

arXiv:2608.04847v1 Announce Type: new Abstract: Despite its cultural relevance and diffusion, queer slang remains underrepresented in Natural Language Processing research. Towards addressing this gap,

Do LLMs Know What Is Private Internally? Probing and Steering Contextual Privacy Norms in Large Language Model Representations

ResearchDGX agent

arXiv:2604.00209v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly deployed in high-stakes settings, yet they frequently violate contextual privacy by disclosing private

Docs: US data labeling companies, like Surge AI and Mercor, that sell training datasets to US AI labs and the government are also selling them to Chinese labs (Anna Tong/Forbes)

HardwareDGX agent

Anna Tong / Forbes: Docs: US data labeling companies, like Surge AI and Mercor, that sell training datasets to US AI labs and the government are also selling them to Chinese labs — The same Silicon Va

Document Optimization for Black-Box Retrieval via Reinforcement Learning

ResearchDGX agent

arXiv:2604.05087v3 Announce Type: replace Abstract: Document expansion is a classical technique for improving retrieval quality, and is attractive since it shifts computation offline, avoiding additio

Does Out-of-Sight Equal Out-of-Mind in CoT Monitorability?

ResearchDGX agent

arXiv:2608.04928v1 Announce Type: new Abstract: Chain-of-thought (CoT) reasoning offers a window into the decision-making of large language models (LLMs), which can be monitored for target behaviors b

DreamWAM: Beyond RGB Future Prediction for World Action Models

Model ReleasesDGX agent

arXiv:2608.04996v1 Announce Type: new Abstract: World Action Models (WAMs) learn action-relevant representations by predicting how the observed world will evolve. Most existing WAMs define this future

DSeq-JEPA: Discriminative Sequential Joint-Embedding Predictive Architecture

ResearchDGX agent

arXiv:2511.17354v4 Announce Type: replace Abstract: Recent advances in self-supervised visual representation learning have demonstrated the effectiveness of predictive latent-space objectives for lear

Dual 3090 setup: 400 pp t/s to 1600 pp t/s on Qwen 3.6 27B... with slightly lower tps.

Model ReleasesDGX agent

First of all, my setup: Ryzen 9 5950x DDR4 3200Mhz 64gb (2x32) Dual 3090s, no NVLINK Runtime: llama.cpp Nvidia Drivers 610 Windows 11 25H2 Qwen 3.6 27B Q8 I've been using llama-server with --split-mod

DXC partners with Primary on zero-trust security for enterprise AI

SafetyDGX agent

DXC Technology Co. today announced a partnership with security startup Primary that makes the information technology services company the exclusive managed services partner for Primary’s zero-trust pl

Dynamic Jailbreaking Attack

Model ReleasesDGX agent

arXiv:2510.02422v4 Announce Type: replace-cross Abstract: Existing gradient-based jailbreak attacks typically optimize a fixed-length adversarial suffix toward a predefined target response with a stat

Dynamical Lie Algebras Cannot Describe Shallow QAOA: Cragged Terrains, Barren Plateaus, and Empirical Hardness Models

ResearchDGX agent

arXiv:2608.04252v1 Announce Type: cross Abstract: The dynamical Lie algebraic (DLA) theory of variational quantum algorithms (VQAs) predicts commonplace exponentially vanishing loss and gradient varia

E^2M: Double Bounded alpha-Divergence Optimization for Tensor-based Discrete Density Estimation

ResearchDGX agent

arXiv:2405.18220v4 Announce Type: replace-cross Abstract: Tensor-based discrete density estimation requires flexible modeling and proper divergence criteria to enable effective learning; however, trad

EA-Graph: Artifact-Anchored Verification Memory for Coding Agents under Upstream Drift

ResearchDGX agent

arXiv:2608.04278v1 Announce Type: cross Abstract: Coding agents increasingly work across sessions, but prose notes can preserve a conclusion without the program state that supported it. After an upstr

Easy to Complete, Hard to Choose: Investigating LLM Performance on the ProverbIT Benchmark

Model ReleasesDGX agent

arXiv:2608.04670v1 Announce Type: cross Abstract: Large Language Models (LLMs) have transformed computational linguistics and achieved remarkable performance across numerous natural language processin

EASy: Towards Efficient LLM-Based Agentic System

AgentsDGX agent

arXiv:2608.04588v1 Announce Type: cross Abstract: Agentic systems have emerged as a promising paradigm for solving complex tasks by coordinating specialized LLM-based agents. However, most existing sy

Echo Flow Networks

Model ReleasesDGX agent

arXiv:2509.24122v3 Announce Type: replace Abstract: At the heart of time-series forecasting (TSF) lies a fundamental challenge: how can models efficiently and effectively capture long-range temporal d

EDATracer: An Agentic Framework for Large-Scale EDA Artifact Analysis

Model ReleasesDGX agent

arXiv:2608.04032v1 Announce Type: cross Abstract: Modern chip design relies on electronic design automation (EDA) tools that generate large, heterogeneous artifacts, including source files, scripts, l

EdgeLM: Edge Demonstrations for Language Models' Table Understanding

ResearchDGX agent

arXiv:2608.04390v1 Announce Type: new Abstract: Large language models (LLMs) perform table-centric prediction through in-context learning, making demonstration selection critical to performance. Exist

Efficient Online Lexicographic Generalized Low-Rank Matrix Bandits

Model ReleasesDGX agent

arXiv:2608.04324v1 Announce Type: cross Abstract: This paper studies generalized low-rank matrix bandits with multiple prioritized objectives. At each round, the learner selects a matrix-valued arm an

EgoAfford: Task-Oriented Affordance Grounding via Egocentric Referring Segmentation

Model ReleasesDGX agent

arXiv:2608.04533v1 Announce Type: new Abstract: Part-level affordance grounding has advanced the localization of functional object regions associated with elemental actions. Extending this capability

Eigenius: A Typed Knowledge-Graph DBMS with Epistemic Stratification and Institution-Mediated Reasoning

AgentsDGX agent

arXiv:2608.04457v1 Announce Type: cross Abstract: As 'AI Scientists' emerge to drive research via the Model Context Protocol (MCP), systems relying on ephemeral scripts will fail. The sheer scale of s

Elbow-Based MoE Routing: A Training-Free Inference Time Plugin for Expert Selection

ResearchDGX agent

arXiv:2608.04401v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) models enable model scaling while maintaining low inference-time compute by activating only a subset of experts per token. Howe

Eliciting Intrinsic Hallucinations in LLMs via Semantically Equivalent Adversarial Attacks

Model ReleasesDGX agent

arXiv:2608.04286v1 Announce Type: new Abstract: Large language models (LLMs) are often used in conjunction with external knowledge sources to improve their factual accuracy and decrease hallucinations

Embedding Large Language Models into Flow Controls: An Agentic Framework for Adaptive and Trustworthy Automated Cooking

AgentsDGX agent

arXiv:2608.04768v1 Announce Type: new Abstract: Automated cooking robots have traditionally relied on predefined procedures and rule-based control, ensuring stable execution but offering limited perso

Emergence of Hierarchical Emotion Organization in Large Language Models

ResearchDGX agent

arXiv:2507.10599v3 Announce Type: replace-cross Abstract: As large language models (LLMs) increasingly power conversational agents, understanding how they model users' emotional states is critical for

EmpaAva: An Open-source Agentic 3D-Avatar Empathetic Live Chatbot

AgentsDGX agent

arXiv:2608.04709v1 Announce Type: new Abstract: This paper presents EmpaAva, to our knowledge the first open-source, agentic 3D-avatar empathetic chatbot, which carries empathetic response generation

Enabling Urgency-aware Robot Swarm Intralogistics using Smart IoT Tags

SafetyDGX agent

arXiv:2608.04721v1 Announce Type: new Abstract: Warehouse items differ in how urgently they must be moved: perishable goods, pharmaceutical shipments, and just-in-time production materials must be del

EndoVLM: An Endoscopy Vision-Language Pre-training Model via Anatomy-Guided Sparsity and Progressive Alignment

SafetyDGX agent

arXiv:2608.04472v1 Announce Type: cross Abstract: The development of foundation models (FMs) is crucial for advancing endoscopic image analysis. However, existing endoscopy FMs mainly rely on self-sup

Energy- and Memory-Efficient PEFT Methods for Personalized On-Device SLMs on Consumer GPUs

Model ReleasesDGX agent

arXiv:2608.04488v1 Announce Type: new Abstract: Despite rapid advances in large language models (LLMs), deploying and personalizing them on resource-constrained devices remains impractical due to high

Energy-Tweedie: Score meets Score, Energy meets Energy

Model ReleasesDGX agent

arXiv:2512.23818v2 Announce Type: replace-cross Abstract: Denoising and score estimation are classically linked through Tweedie's formula, which relates the posterior mean under Gaussian noise to the

Enforcing data residency with single-Region Claude Code on Amazon Bedrock

Model ReleasesDGX agent

A regulated customer needed all Claude Code inference processed in a single AWS Region (London), not just in-geography. This post shows two ways to pin Claude Code on Amazon Bedrock to one Region: an

Enhancing Low Back Pain Assessment with Diffusion Models for Lumbar Spine MRI Segmentation

ResearchDGX agent

arXiv:2608.04906v1 Announce Type: new Abstract: This study introduces a diffusion-based framework for robust and accurate semantic segmentation of lumbar spine MRI scans from patients with low back pa

Enhancing Trustworthy Clinical Diagnosis Decision-Making in Large Language Models via Etiology-Aware Attention Supervision

Model ReleasesDGX agent

arXiv:2508.00285v2 Announce Type: replace Abstract: Objective: Large Language Models (LLMs) have demonstrated strong capabilities in medical text understanding and generation. However, their trustwort

Equitable System-Prompt Selection via Constrained Mixed-Strategy GroupDRO

ApplicationsDGX agent

arXiv:2608.04339v1 Announce Type: cross Abstract: Large language models are increasingly used for information seeking, yet semantically equivalent questions phrased in different ways can receive answe

EuroExec: Frontier Language Models Fall Short of Expert Judgment on European Executive Decision Tasks

Model ReleasesDGX agent

arXiv:2608.04549v1 Announce Type: cross Abstract: Frontier LLMs are increasingly put to use on open-ended complex questions, different in nature from the ones they are typically evaluated on. We dedic

Evaluating the Diagnostic Robustness of Vision-Language Models Under Visual and Textual Perturbations

SafetyDGX agent

arXiv:2608.04885v1 Announce Type: cross Abstract: Standard accuracy metrics for VLMs often mask significant reliability failures in sensitive domains. In this work, we utilize a histopathology-validat

Evaluating Theory of Mind in Reasoning Models: Robustness over Reasoning

ResearchDGX agent

arXiv:2608.04646v1 Announce Type: new Abstract: Large language models (LLMs) have recently shown strong performance on Theory of Mind (ToM) tests, prompting debate about the nature and validity of the

Evaluation Pitfalls and Sparsity Limitations in LLM-based Confidence Estimates for Classification

ResearchDGX agent

arXiv:2608.04899v1 Announce Type: new Abstract: Confidence estimation is essential when LLMs are used for classification, indicating when predictions can be trusted. However, common approaches such as

EviGraph: Evidence-Guided Autonomous Research Agents

AgentsDGX agent

arXiv:2608.04738v1 Announce Type: new Abstract: Autonomous research agents can generate hypotheses, execute experiments, and draft manuscripts, yet their outputs often contain unsupported claims and i

EvolveNet: Collaborative Harness Evolution for Agent Self-Improvement

Local AiDGX agent

arXiv:2608.04968v1 Announce Type: new Abstract: The capabilities of an LLM agent depend not only on its model but on the harness: the executable program that constructs context, invokes tools, verifie

EvtGraph: Event-Adaptive Compression for Sparse Temporal Graph Learning in Multimodal Time Series

ResearchDGX agent

arXiv:2608.04368v1 Announce Type: new Abstract: Multimodal temporal data are inherently irregular and uneven in information density, yet most models rely on uniform discretization, leading to ineffici

← Previous
1…8283848586…1408
Next →