AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
Human
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
84,531 results
6 Aug 2026

Optimal Training-Time Scaling in Gradual Adaptation

ResearchDGX agent

arXiv:2608.04927v1 Announce Type: new Abstract: In gradual adaptation, how should the training time on each task change as the number of intermediate tasks increases? We study this question for overpa

Optimizing the Preconditioner: A Black-box Online-to-Nonconvex Conversion with Static Regret Minimization Oracles

ResearchDGX agent

arXiv:2607.17607v2 Announce Type: replace Abstract: We study whether stochastic nonconvex optimization can be reduced to ordinary static regret minimization in online convex optimization in a black-bo

Optimizing What Policies Learn From: Recoverability-aware Rollout Intervention Learning

SafetyDGX agent

arXiv:2608.05080v1 Announce Type: cross Abstract: Critic-free group-based reinforcement learning has become a scalable approach for post-training large language models. However, most existing methods

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

ORACLE: A Multi-Objective Reinforcement Learning-Based Analog Circuit Design Optimizer with Large Language Models-Guided Exploration

ResearchDGX agent

arXiv:2608.04999v1 Announce Type: cross Abstract: Analog circuit design automation using reinforcement learning (RL) has emerged as a promising approach for reducing manual effort. However, many exist

Our Black Hat talk on the OpenAI-Hugging Face incident is now live on youtube. This is a watershed moment for the industry. I encourage all …

ToolsDGX agent

Our Black Hat talk on the OpenAI-Hugging Face incident is now live on youtube. This is a watershed moment for the industry. I encourage all defenders to watch, consider how attack dynamics will immine

Out-Of-The-Loop Multi-Fidelity Bayesian Optimization

ApplicationsDGX agent

arXiv:2608.04113v1 Announce Type: cross Abstract: Black-box optimization is a ubiquitous problem in science and engineering, often dealing with expensive objective functions with cheaper lower-fidelit

OutLangSplat: 3D Language Gaussian Splatting for UAV Outdoor Scenes

SafetyDGX agent

arXiv:2608.04560v1 Announce Type: new Abstract: 3D Language Gaussian Splatting embeds open-vocabulary language features into 3D Gaussian Splatting, providing an efficient explicit representation for t

Outlook where frontier AI is headed next 18 months: The AI reasoning training + harness loop works if you can produce enough data and reason…

AgentsDGX agent

Outlook where frontier AI is headed next 18 months: The AI reasoning training + harness loop works if you can produce enough data and reasoning traces (via verifiers). Proven with code and math result

Overcoming Statistical Bias in Action-Controllable World Models

SafetyDGX agent

arXiv:2608.04653v1 Announce Type: new Abstract: Action-conditioned world models aim to predict how visual environments evolve under an agent's actions. Yet future frames are often highly predictable f

PADFormer: Pose-agnostic Anomaly Detection from Sparse View Images

Model ReleasesDGX agent

arXiv:2608.04210v1 Announce Type: new Abstract: Pose-agnostic Anomaly Detection (PAD) remains challenging as anomalies can appear under arbitrary viewpoints, requiring methods to handle significant po

Patients-like-me: A Variational LM--GNN Framework for Explainable Clinical Prediction

Local AiDGX agent

arXiv:2608.04193v1 Announce Type: cross Abstract: Language models (LMs) offer strong textual representations for electronic health records (EHRs), but they encode patient sequences in isolation and pr

Perception Before Reasoning: Dynamic Latent Reasoning for Video Understanding and Question Answering

Local AiDGX agent

arXiv:2608.04124v1 Announce Type: cross Abstract: Video question answering requires models to ground language queries in visual evidence and, when necessary, reason over that evidence across time. Exi

Persistent Object Narratives for Token-Efficient Video Language Models

Model ReleasesDGX agent

arXiv:2608.04866v1 Announce Type: new Abstract: Video large language models (Video-LLMs) have made strong progress in open-ended video understanding. However, their visual interfaces remain token-inte

Personalized Federated Sparse Adaptation of Time-Series Foundation Models

Model ReleasesDGX agent

arXiv:2608.04695v1 Announce Type: cross Abstract: Federated adaptation of time-series foundation models (TSFMs) is attractive for building energy forecasting because meter data are private, distribute

Physics-informed reduced-order modelling with equivariant spectral submanifolds

Model ReleasesDGX agent

arXiv:2608.04239v1 Announce Type: new Abstract: Spectral submanifold (SSM) reduction has emerged as a mathematically principled route to reliable nonlinear reduced-order models, capturing dynamics bey

PhysMind: From Video to Executable Worlds for Training-Free Physical Reasoning

Model ReleasesDGX agent

arXiv:2608.04575v1 Announce Type: cross Abstract: Reliable physical reasoning from video requires understanding how objects move, interact, and respond to interventions. Existing vision-language model

PICopilot: An LLM-based Agentic Framework for Assisting Photonic Integrated Circuit Design via Script Generation

Model ReleasesDGX agent

arXiv:2608.01791v2 Announce Type: replace-cross Abstract: The rapid development of photonic integrated circuits (PICs) is shifting the design flow from traditional graphical user interface (GUI)-based

PitchBook: AI voice startups raised 7B in Q1 2026, up from 1B in Q1 2025, as OpenAI and Google bet on voice as the main interface for next-gen AI agents (Cristina Criddle/Financial Times)

IndustryDGX agent

Cristina Criddle / Financial Times: PitchBook: AI voice startups raised 7B in Q1 2026, up from 1B in Q1 2025, as OpenAI and Google bet on voice as the main interface for next-gen AI agents — OpenAI an

Place-it-R1: Unlocking Environment-aware Reasoning Potential of MLLM for Video Object Insertion

Local AiDGX agent

arXiv:2603.06140v2 Announce Type: replace-cross Abstract: Video object insertion is fundamental to video editing, yet existing diffusion methods often produce visually plausible but physically inconsi

Plus and Pro users also now have a slider to choose how much reasoning effort ChatGPT puts into each response. We think it’s easier to use, …

Model ReleasesDGX agent

OpenAI announced that Plus and Pro subscribers now have a slider to adjust the amount of reasoning effort ChatGPT applies to each response. The update employs GPT‑5.6 Sol for both Instant and deep rea

Plus and Pro users can access the updated version of GPT‑5.6 Sol in ChatGPT along with the new slider starting today. This version of GPT-5.…

Model ReleasesDGX agent

Plus and Pro users can access the updated version of GPT‑5.6 Sol in ChatGPT along with the new slider starting today. This version of GPT-5.6 Sol is for everyday chats, so it will only be available in

Poly-OPD: Heterogeneous Multi-Teacher On-Policy Distillation for Capability-Selectable Flow Models

SafetyDGX agent

arXiv:2608.04349v1 Announce Type: new Abstract: Leading open text-to-image models often carry complementary strengths: one may lead on preference-aligned aesthetics while another follows compositional

Predict, Then Retrieve: Cross-Instance Future-State Retrieval from Video Prefixes

Model ReleasesDGX agent

arXiv:2608.04426v1 Announce Type: cross Abstract: We introduce Predictive State Retrieval (PSR), a task in which a model observes a short video prefix and a temporal question about an object's future

Predicting Brain Morphometry with MT-GNN: Mesh Evolution in Continuous Time with Graph-Based Metric Tensor Embeddings

ResearchDGX agent

arXiv:2608.05132v1 Announce Type: new Abstract: Predicting how a subcortical structure's shape will evolve from a few prior scans could support prognosis and clinical-trial enrichment. Existing longit

Preverbal Uninflected and Underived Roots in Mapudungun. Wuno and Its Implications

ResearchDGX agent

arXiv:2608.04869v1 Announce Type: new Abstract: This study examines the grammatical status of preverbal uninflected and underived roots in Mapudungun, with particular focus on wuno 'return/re-'. Throu

PriDyG: Privacy-preserving Dynamic Graph Inference with LLM-GNN Collaboration

ResearchDGX agent

arXiv:2608.04255v1 Announce Type: cross Abstract: Graph inference over relational data can expose sensitive edge information, and this risk becomes more severe in dynamic graphs, where repeated model

PRIMAL3: Pathfinding via Reinforcement and Imitation Multi-Agent Learning - Leveraging LaCAM3

SafetyDGX agent

arXiv:2608.04905v1 Announce Type: new Abstract: We present PRIMAL3, an ultra-large-scale learning-based framework for multi-agent pathfinding (MAPF) that integrates reinforcement learning, topology-aw

Privacy-Preserving Action Recognition: Taxonomy, Methods, and Privacy-Utility Trade-offs

Model ReleasesDGX agent

arXiv:2608.04501v1 Announce Type: new Abstract: Video surveillance in public safety, healthcare, and smart environments has made continuous human monitoring routine, raising real risks to personal ide

Privileged, but Biased: How PI-Conditioned Teachers Break Self-Distillation

SafetyDGX agent

arXiv:2608.04794v1 Announce Type: new Abstract: Self-distillation (SD) has emerged as a compute-efficient alternative to reinforcement learning with verifiable rewards: a self-teacher, conditioned on

Promptable Animal Pose Tracking Across Species

Local AiDGX agent

arXiv:2608.04995v1 Announce Type: new Abstract: Animal pose estimation and tracking is important for wildlife monitoring and conservation research, and with limited expert time for labelling automated

Protoreasoning in Tiny Transformers

Model ReleasesDGX agent

arXiv:2608.04980v1 Announce Type: cross Abstract: We show that tiny transformers can profitably employ a simple form of Chain of Thought, which we call protoreasoning, allowing us to study step-by-ste

Prototype-based Self-Supervised Multimodal Learning for PPG and Accelerometry Signals

ResearchDGX agent

arXiv:2510.09764v2 Announce Type: replace Abstract: Modeling multi-modal time-series data is critical for capturing system-level dynamics, particularly in biosignals where modalities such as ECG, PPG,

Provable Limits and Certified Deferral for Verbalized Uncertainty in Small Language Models

Local AiDGX agent

arXiv:2608.05064v1 Announce Type: cross Abstract: Small open-weight language models increasingly run in private, offline, and cost-sensitive settings, where the key deployment question is not only wha

PSI3D: Plug-and-Play 3D Stochastic Inference with Slice-wise Latent Diffusion Prior

ResearchDGX agent

arXiv:2512.18367v2 Announce Type: replace-cross Abstract: Diffusion models are highly expressive image priors for Bayesian inverse problems. However, most diffusion models cannot operate on large-scal

Pun Intended: Multi-Agent Translation of Wordplay with Contrastive Learning and Phonetic-Semantic Embeddings for CLEF JOKER 2025 Task 2

AgentsDGX agent

arXiv:2507.06506v2 Announce Type: replace-cross Abstract: Translating wordplay across languages presents unique challenges that have long confounded both professional human translators and machine tra

PURPOSE: Poisoning Conflict Resolution in RAG via Proxy-Fact-Grounded Updates

ResearchDGX agent

arXiv:2608.04756v1 Announce Type: cross Abstract: In Retrieval-Augmented Generation (RAG), post-retrieval conflict resolution arbitrates among noisy or contradictory retrieved passages. However, the r

Q-CueGraph: Query-Conditioned Visual Evidence Graphs for Multimodal Reasoning

SafetyDGX agent

arXiv:2608.04452v1 Announce Type: cross Abstract: High-resolution pixels and crop or zoom tools give multimodal large language models the ability to inspect an image, but they do not provide a reliabl

Radar4D-VLM: Proposal-Grounded Temporal 4D Radar Reasoning Across Frozen Language Models

Model ReleasesDGX agent

arXiv:2608.04130v1 Announce Type: new Abstract: Vision-language models for autonomous driving primarily rely on cameras and LiDAR, leaving 4D radar largely unexplored as a standalone perceptual modali

RAG-Stack: Co-Optimizing RAG Serving Performance and Quality

ResearchDGX agent

arXiv:2608.03487v1 Announce Type: cross Abstract: Retrieval-augmented generation (RAG), which augments large language model (LLM) generation with information retrieved from databases, has become a wid

Random features for Grassmannian kernel approximation with bounded rank-one projections

ResearchDGX agent

arXiv:2608.04227v1 Announce Type: new Abstract: We propose a family of random feature maps for scalable kernel machines on low-dimensional subspaces, ie on the Grassmannian manifold. Such representati

Reachability in 3-VAS

ResearchDGX agent

arXiv:2608.04786v1 Announce Type: new Abstract: We settle the exact complexity of the reachability problem in (stateless) vector addition systems (VAS) in fixed low dimension. In dimensions 2-4 it has

Reading Between the Frames: Interpreting Implicit and Non-literal Meaning in Social Media Videos

Model ReleasesDGX agent

arXiv:2608.04939v1 Announce Type: new Abstract: Social media videos often communicate meanings that go beyond their visible actions, captions, or speech. A mundane clip may become humorous, ironic, or

Real-time probabilistic tsunami forecasting via generative AI

SafetyDGX agent

arXiv:2608.04327v1 Announce Type: new Abstract: Explicit onshore tsunami inundation forecasting can improve public risk awareness, but deterministically predicted inundation boundaries under highly un

Reasoning Core: Designing Broad Procedural Data for Completion-Supervised Reasoning Training

ResearchDGX agent

arXiv:2608.05148v1 Announce Type: new Abstract: Procedural generators produce useful verifiable reasoning problems at scale, but have received less attention as data for completion-supervised fine-tun

ReCodeAgent: A Multi-agent Workflow for Language-Agnostic Translation and Validation of Large-Scale Repositories

AgentsDGX agent

arXiv:2604.07341v2 Announce Type: replace-cross Abstract: Most repository-level code translation and validation techniques have been evaluated on a single source-target programming language (PL) pair,

Reconstructing Persistent Worlds from Narratives for Narrative-Grounded Interactive Experiences

ResearchDGX agent

arXiv:2608.04037v1 Announce Type: cross Abstract: Designing narrative-grounded interactive experiences remains labor-intensive because interactive content must align with the underlying world implied

Recurrent Residual Quantization: A Progressive Multi-Precision Representation for LLMs

ResearchDGX agent

arXiv:2608.04048v1 Announce Type: cross Abstract: Serving large language models (LLMs) under diverse deployment constraints requires flexible trade-offs between accuracy, memory footprint, and through

RegisterBridgeMM: A Register-Centric Framework for RGB-Infrared Object Detection

ResearchDGX agent

arXiv:2608.04833v1 Announce Type: new Abstract: RGB-infrared (RGB-IR) object detection benefits from complementary visible and thermal cues, but effective fusion remains challenging under illumination

ReGround: Restoring Visual Grounding in Multi-Step Reasoning through Self-Diagnosis and Visual Re-Examination

Model ReleasesDGX agent

arXiv:2608.04385v1 Announce Type: new Abstract: Vision-Language Models (VLMs) often lose visual grounding during multi-step reasoning: as reasoning chains grow longer, later inference steps rely incre

Regularization can make diffusion models more efficient

ResearchDGX agent

arXiv:2502.09151v3 Announce Type: replace Abstract: Diffusion models are one of the key architectures of generative AI. Their main drawback, however, is the computational costs. This study indicates t

Reinforcement Learning and Consumption-Savings Behavior

ResearchDGX agent

arXiv:2510.20748v2 Announce Type: replace-cross Abstract: This paper demonstrates how reinforcement learning can explain two puzzling empirical patterns in household consumption behavior during econom

Relational Response Fields: A General Theory of Black-Box LLM Response Consistency and Recovery

ResearchDGX agent

arXiv:2608.04552v1 Announce Type: new Abstract: Black-box language-model reliability is commonly pursued by sampling, prompting, voting, verifying, or iteratively revising individual answers. We ask a

Relevant but Incomplete: Referential Dangling as a Paradigm-Level Failure Mode in Hard Prompt Compression

Model ReleasesDGX agent

arXiv:2608.04569v1 Announce Type: new Abstract: Hard prompt compression reduces long-context inference cost by independently scoring tokens, sentences, or chunks and retaining the highest-scoring unit

RepairFormer: Automated Repair of Structured Inputs Using Transformers

Model ReleasesDGX agent

arXiv:2608.05060v1 Announce Type: cross Abstract: Structured input files such as JSON, DOT, OBJ, INI, S-expression, and TinyC are widely used in software systems, but small corruptions can cause parse

RepoProbe: Benchmarking Architecture-Aware Repository Comprehension with Checklists

Model ReleasesDGX agent

arXiv:2608.04783v1 Announce Type: cross Abstract: The integration of Large Language Models (LLMs) into software engineering has shifted the focus from function-level generation to repository-scale ass

Report: OpenAI’s upcoming AI speaker will be shaped like a donut and cost around $300

IndustryDGX agent

More details have emerged from Bloomberg about OpenAI Group PBC’s long-awaited hardware device, which was previously described as an artificial intelligence-enabled speaker that will be a “physical ma

Representational separation between unitary and channel quantum generative models via shared classical randomness at shallow depth

Local AiDGX agent

arXiv:2608.05110v1 Announce Type: cross Abstract: Near-term quantum hardware limits circuit depth and often imposes geometrically local connectivity for quantum generative models, restricting the outp

Representing Visual Evidence for Item Difficulty Prediction: Visual Textualization and Image-Native Modeling

ResearchDGX agent

arXiv:2608.04554v1 Announce Type: new Abstract: Predicting item difficulty from content can provide an initial estimate for newly developed questions before sufficient student responses are available.

RESPClinBench: Benchmarking Multimodal Clinical Decision-Making and Longitudinal Disease Management in Respiratory Specialty Care

Model ReleasesDGX agent

arXiv:2608.04514v1 Announce Type: new Abstract: Background: Respiratory specialty care requires multimodal interpretation, longitudinal risk assessment, guideline-concordant intervention, and whole-co

ResPlan: A Large-Scale Vector-Graph Dataset of 17,000 Residential Floor Plans

Model ReleasesDGX agent

arXiv:2508.14006v2 Announce Type: replace Abstract: We introduce ResPlan, a dataset of 17,000 residential floor plans with vector geometry, room-connectivity graphs, and metric-scale coordinates. Each

← Previous
1…8889909192…1409
Next →