AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
84,460 results
6 Aug 2026

Persistent Object Narratives for Token-Efficient Video Language Models

Model ReleasesDGX agent

arXiv:2608.04866v1 Announce Type: new Abstract: Video large language models (Video-LLMs) have made strong progress in open-ended video understanding. However, their visual interfaces remain token-inte

Personalized Federated Sparse Adaptation of Time-Series Foundation Models

Model ReleasesDGX agent

arXiv:2608.04695v1 Announce Type: cross Abstract: Federated adaptation of time-series foundation models (TSFMs) is attractive for building energy forecasting because meter data are private, distribute

Physics-informed reduced-order modelling with equivariant spectral submanifolds

Model ReleasesDGX agent

arXiv:2608.04239v1 Announce Type: new Abstract: Spectral submanifold (SSM) reduction has emerged as a mathematically principled route to reliable nonlinear reduced-order models, capturing dynamics bey

Content type
AllBlogX PostPaperYouTubeRedditGitHub

PhysMind: From Video to Executable Worlds for Training-Free Physical Reasoning

Model ReleasesDGX agent

arXiv:2608.04575v1 Announce Type: cross Abstract: Reliable physical reasoning from video requires understanding how objects move, interact, and respond to interventions. Existing vision-language model

PICopilot: An LLM-based Agentic Framework for Assisting Photonic Integrated Circuit Design via Script Generation

Model ReleasesDGX agent

arXiv:2608.01791v2 Announce Type: replace-cross Abstract: The rapid development of photonic integrated circuits (PICs) is shifting the design flow from traditional graphical user interface (GUI)-based

PitchBook: AI voice startups raised 7B in Q1 2026, up from 1B in Q1 2025, as OpenAI and Google bet on voice as the main interface for next-gen AI agents (Cristina Criddle/Financial Times)

IndustryDGX agent

Cristina Criddle / Financial Times: PitchBook: AI voice startups raised 7B in Q1 2026, up from 1B in Q1 2025, as OpenAI and Google bet on voice as the main interface for next-gen AI agents — OpenAI an

Place-it-R1: Unlocking Environment-aware Reasoning Potential of MLLM for Video Object Insertion

Local AiDGX agent

arXiv:2603.06140v2 Announce Type: replace-cross Abstract: Video object insertion is fundamental to video editing, yet existing diffusion methods often produce visually plausible but physically inconsi

Plus and Pro users also now have a slider to choose how much reasoning effort ChatGPT puts into each response. We think it’s easier to use, …

Model ReleasesDGX agent

OpenAI announced that Plus and Pro subscribers now have a slider to adjust the amount of reasoning effort ChatGPT applies to each response. The update employs GPT‑5.6 Sol for both Instant and deep rea

Plus and Pro users can access the updated version of GPT‑5.6 Sol in ChatGPT along with the new slider starting today. This version of GPT-5.…

Model ReleasesDGX agent

Plus and Pro users can access the updated version of GPT‑5.6 Sol in ChatGPT along with the new slider starting today. This version of GPT-5.6 Sol is for everyday chats, so it will only be available in

Poly-OPD: Heterogeneous Multi-Teacher On-Policy Distillation for Capability-Selectable Flow Models

SafetyDGX agent

arXiv:2608.04349v1 Announce Type: new Abstract: Leading open text-to-image models often carry complementary strengths: one may lead on preference-aligned aesthetics while another follows compositional

Predict, Then Retrieve: Cross-Instance Future-State Retrieval from Video Prefixes

Model ReleasesDGX agent

arXiv:2608.04426v1 Announce Type: cross Abstract: We introduce Predictive State Retrieval (PSR), a task in which a model observes a short video prefix and a temporal question about an object's future

Predicting Brain Morphometry with MT-GNN: Mesh Evolution in Continuous Time with Graph-Based Metric Tensor Embeddings

ResearchDGX agent

arXiv:2608.05132v1 Announce Type: new Abstract: Predicting how a subcortical structure's shape will evolve from a few prior scans could support prognosis and clinical-trial enrichment. Existing longit

Preverbal Uninflected and Underived Roots in Mapudungun. Wuno and Its Implications

ResearchDGX agent

arXiv:2608.04869v1 Announce Type: new Abstract: This study examines the grammatical status of preverbal uninflected and underived roots in Mapudungun, with particular focus on wuno 'return/re-'. Throu

PriDyG: Privacy-preserving Dynamic Graph Inference with LLM-GNN Collaboration

ResearchDGX agent

arXiv:2608.04255v1 Announce Type: cross Abstract: Graph inference over relational data can expose sensitive edge information, and this risk becomes more severe in dynamic graphs, where repeated model

PRIMAL3: Pathfinding via Reinforcement and Imitation Multi-Agent Learning - Leveraging LaCAM3

SafetyDGX agent

arXiv:2608.04905v1 Announce Type: new Abstract: We present PRIMAL3, an ultra-large-scale learning-based framework for multi-agent pathfinding (MAPF) that integrates reinforcement learning, topology-aw

Privacy-Preserving Action Recognition: Taxonomy, Methods, and Privacy-Utility Trade-offs

Model ReleasesDGX agent

arXiv:2608.04501v1 Announce Type: new Abstract: Video surveillance in public safety, healthcare, and smart environments has made continuous human monitoring routine, raising real risks to personal ide

Privileged, but Biased: How PI-Conditioned Teachers Break Self-Distillation

SafetyDGX agent

arXiv:2608.04794v1 Announce Type: new Abstract: Self-distillation (SD) has emerged as a compute-efficient alternative to reinforcement learning with verifiable rewards: a self-teacher, conditioned on

Promptable Animal Pose Tracking Across Species

Local AiDGX agent

arXiv:2608.04995v1 Announce Type: new Abstract: Animal pose estimation and tracking is important for wildlife monitoring and conservation research, and with limited expert time for labelling automated

Protoreasoning in Tiny Transformers

Model ReleasesDGX agent

arXiv:2608.04980v1 Announce Type: cross Abstract: We show that tiny transformers can profitably employ a simple form of Chain of Thought, which we call protoreasoning, allowing us to study step-by-ste

Prototype-based Self-Supervised Multimodal Learning for PPG and Accelerometry Signals

ResearchDGX agent

arXiv:2510.09764v2 Announce Type: replace Abstract: Modeling multi-modal time-series data is critical for capturing system-level dynamics, particularly in biosignals where modalities such as ECG, PPG,

Provable Limits and Certified Deferral for Verbalized Uncertainty in Small Language Models

Local AiDGX agent

arXiv:2608.05064v1 Announce Type: cross Abstract: Small open-weight language models increasingly run in private, offline, and cost-sensitive settings, where the key deployment question is not only wha

PSI3D: Plug-and-Play 3D Stochastic Inference with Slice-wise Latent Diffusion Prior

ResearchDGX agent

arXiv:2512.18367v2 Announce Type: replace-cross Abstract: Diffusion models are highly expressive image priors for Bayesian inverse problems. However, most diffusion models cannot operate on large-scal

Pun Intended: Multi-Agent Translation of Wordplay with Contrastive Learning and Phonetic-Semantic Embeddings for CLEF JOKER 2025 Task 2

AgentsDGX agent

arXiv:2507.06506v2 Announce Type: replace-cross Abstract: Translating wordplay across languages presents unique challenges that have long confounded both professional human translators and machine tra

PURPOSE: Poisoning Conflict Resolution in RAG via Proxy-Fact-Grounded Updates

ResearchDGX agent

arXiv:2608.04756v1 Announce Type: cross Abstract: In Retrieval-Augmented Generation (RAG), post-retrieval conflict resolution arbitrates among noisy or contradictory retrieved passages. However, the r

Q-CueGraph: Query-Conditioned Visual Evidence Graphs for Multimodal Reasoning

SafetyDGX agent

arXiv:2608.04452v1 Announce Type: cross Abstract: High-resolution pixels and crop or zoom tools give multimodal large language models the ability to inspect an image, but they do not provide a reliabl

Radar4D-VLM: Proposal-Grounded Temporal 4D Radar Reasoning Across Frozen Language Models

Model ReleasesDGX agent

arXiv:2608.04130v1 Announce Type: new Abstract: Vision-language models for autonomous driving primarily rely on cameras and LiDAR, leaving 4D radar largely unexplored as a standalone perceptual modali

RAG-Stack: Co-Optimizing RAG Serving Performance and Quality

ResearchDGX agent

arXiv:2608.03487v1 Announce Type: cross Abstract: Retrieval-augmented generation (RAG), which augments large language model (LLM) generation with information retrieved from databases, has become a wid

Random features for Grassmannian kernel approximation with bounded rank-one projections

ResearchDGX agent

arXiv:2608.04227v1 Announce Type: new Abstract: We propose a family of random feature maps for scalable kernel machines on low-dimensional subspaces, ie on the Grassmannian manifold. Such representati

Reachability in 3-VAS

ResearchDGX agent

arXiv:2608.04786v1 Announce Type: new Abstract: We settle the exact complexity of the reachability problem in (stateless) vector addition systems (VAS) in fixed low dimension. In dimensions 2-4 it has

Reading Between the Frames: Interpreting Implicit and Non-literal Meaning in Social Media Videos

Model ReleasesDGX agent

arXiv:2608.04939v1 Announce Type: new Abstract: Social media videos often communicate meanings that go beyond their visible actions, captions, or speech. A mundane clip may become humorous, ironic, or

Real-time probabilistic tsunami forecasting via generative AI

SafetyDGX agent

arXiv:2608.04327v1 Announce Type: new Abstract: Explicit onshore tsunami inundation forecasting can improve public risk awareness, but deterministically predicted inundation boundaries under highly un

Reasoning Core: Designing Broad Procedural Data for Completion-Supervised Reasoning Training

ResearchDGX agent

arXiv:2608.05148v1 Announce Type: new Abstract: Procedural generators produce useful verifiable reasoning problems at scale, but have received less attention as data for completion-supervised fine-tun

ReCodeAgent: A Multi-agent Workflow for Language-Agnostic Translation and Validation of Large-Scale Repositories

AgentsDGX agent

arXiv:2604.07341v2 Announce Type: replace-cross Abstract: Most repository-level code translation and validation techniques have been evaluated on a single source-target programming language (PL) pair,

Reconstructing Persistent Worlds from Narratives for Narrative-Grounded Interactive Experiences

ResearchDGX agent

arXiv:2608.04037v1 Announce Type: cross Abstract: Designing narrative-grounded interactive experiences remains labor-intensive because interactive content must align with the underlying world implied

Recurrent Residual Quantization: A Progressive Multi-Precision Representation for LLMs

ResearchDGX agent

arXiv:2608.04048v1 Announce Type: cross Abstract: Serving large language models (LLMs) under diverse deployment constraints requires flexible trade-offs between accuracy, memory footprint, and through

RegisterBridgeMM: A Register-Centric Framework for RGB-Infrared Object Detection

ResearchDGX agent

arXiv:2608.04833v1 Announce Type: new Abstract: RGB-infrared (RGB-IR) object detection benefits from complementary visible and thermal cues, but effective fusion remains challenging under illumination

ReGround: Restoring Visual Grounding in Multi-Step Reasoning through Self-Diagnosis and Visual Re-Examination

Model ReleasesDGX agent

arXiv:2608.04385v1 Announce Type: new Abstract: Vision-Language Models (VLMs) often lose visual grounding during multi-step reasoning: as reasoning chains grow longer, later inference steps rely incre

Regularization can make diffusion models more efficient

ResearchDGX agent

arXiv:2502.09151v3 Announce Type: replace Abstract: Diffusion models are one of the key architectures of generative AI. Their main drawback, however, is the computational costs. This study indicates t

Reinforcement Learning and Consumption-Savings Behavior

ResearchDGX agent

arXiv:2510.20748v2 Announce Type: replace-cross Abstract: This paper demonstrates how reinforcement learning can explain two puzzling empirical patterns in household consumption behavior during econom

Relational Response Fields: A General Theory of Black-Box LLM Response Consistency and Recovery

ResearchDGX agent

arXiv:2608.04552v1 Announce Type: new Abstract: Black-box language-model reliability is commonly pursued by sampling, prompting, voting, verifying, or iteratively revising individual answers. We ask a

Relevant but Incomplete: Referential Dangling as a Paradigm-Level Failure Mode in Hard Prompt Compression

Model ReleasesDGX agent

arXiv:2608.04569v1 Announce Type: new Abstract: Hard prompt compression reduces long-context inference cost by independently scoring tokens, sentences, or chunks and retaining the highest-scoring unit

RepairFormer: Automated Repair of Structured Inputs Using Transformers

Model ReleasesDGX agent

arXiv:2608.05060v1 Announce Type: cross Abstract: Structured input files such as JSON, DOT, OBJ, INI, S-expression, and TinyC are widely used in software systems, but small corruptions can cause parse

RepoProbe: Benchmarking Architecture-Aware Repository Comprehension with Checklists

Model ReleasesDGX agent

arXiv:2608.04783v1 Announce Type: cross Abstract: The integration of Large Language Models (LLMs) into software engineering has shifted the focus from function-level generation to repository-scale ass

Report: OpenAI’s upcoming AI speaker will be shaped like a donut and cost around $300

IndustryDGX agent

More details have emerged from Bloomberg about OpenAI Group PBC’s long-awaited hardware device, which was previously described as an artificial intelligence-enabled speaker that will be a “physical ma

Representational separation between unitary and channel quantum generative models via shared classical randomness at shallow depth

Local AiDGX agent

arXiv:2608.05110v1 Announce Type: cross Abstract: Near-term quantum hardware limits circuit depth and often imposes geometrically local connectivity for quantum generative models, restricting the outp

Representing Visual Evidence for Item Difficulty Prediction: Visual Textualization and Image-Native Modeling

ResearchDGX agent

arXiv:2608.04554v1 Announce Type: new Abstract: Predicting item difficulty from content can provide an initial estimate for newly developed questions before sufficient student responses are available.

RESPClinBench: Benchmarking Multimodal Clinical Decision-Making and Longitudinal Disease Management in Respiratory Specialty Care

Model ReleasesDGX agent

arXiv:2608.04514v1 Announce Type: new Abstract: Background: Respiratory specialty care requires multimodal interpretation, longitudinal risk assessment, guideline-concordant intervention, and whole-co

ResPlan: A Large-Scale Vector-Graph Dataset of 17,000 Residential Floor Plans

Model ReleasesDGX agent

arXiv:2508.14006v2 Announce Type: replace Abstract: We introduce ResPlan, a dataset of 17,000 residential floor plans with vector geometry, room-connectivity graphs, and metric-scale coordinates. Each

Rethinking Pixel Mean Flows via Interval Denoiser

ResearchDGX agent

arXiv:2608.04818v1 Announce Type: new Abstract: Modern diffusion and flow-based models are increasingly moving toward few-step, latent-free generation to bypass the computational overhead of multi-ste

Rethinking Reservoir Pruning: A Dynamical Perspective for Echo State Networks

ApplicationsDGX agent

arXiv:2608.04593v1 Announce Type: cross Abstract: Echo State Networks (ESNs) offer an efficient framework for temporal prediction, but their randomly initialized reservoirs are often over-parameterize

Retrieve in Time, Correct in Frequency

Model ReleasesDGX agent

arXiv:2608.04527v1 Announce Type: new Abstract: Frozen vision-language-action (VLA) policies generate temporally extended action chunks, but long-horizon manipulation remains vulnerable to accumulated

Revealed Rationality: Label-Free Evaluation and Regularization from Representation Theorems

ResearchDGX agent

arXiv:2608.05015v1 Announce Type: cross Abstract: Representation theorems in decision theory establish that behavior satisfies certain axioms if and only if it can be rationalized by a well-defined ob

Review Text as a Leading Indicator of Displayed Reputation in Platform Rating Systems: Evidence from 34 U.S. Short-Term Rental Markets

ResearchDGX agent

arXiv:2504.14053v2 Announce Type: replace-cross Abstract: Rating systems on accommodation platforms suffer from a familiar problem: nearly every listing displays a nearly perfect score, so the number

Revisiting Pose Sensitivity in Splat-based Computed Tomography under Sparse-view Reconstruction

ApplicationsDGX agent

arXiv:2608.04752v1 Announce Type: new Abstract: X-ray computed tomography (CT) reconstructs volumetric representations of objects from projection images obtained by transmitting X-rays through a targe

Reward Structure Shapes the Interaction Between Episodic Exploration and Neural Memory in Reinforcement Learning

SafetyDGX agent

arXiv:2608.05111v1 Announce Type: new Abstract: In partially observable reinforcement learning, agents face a dual bottleneck: they must explore to encounter rewarding states and retain that experienc

REZE: Recognition-Based Zero-Shot Extraction for Video Temporal Grounding

ResearchDGX agent

arXiv:2608.04480v1 Announce Type: new Abstract: Video temporal grounding (VTG) refers to the task of identifying the time interval in a video that corresponds to a given natural-language query. A comm

RiboSphere: Learning Unified and Efficient Representations of RNA Structures

ResearchDGX agent

arXiv:2603.19636v2 Announce Type: replace Abstract: Accurate RNA structure modeling remains difficult because RNA backbones are highly flexible, non-canonical interactions are prevalent, and experimen

Right Reset: Chunking by Prefix Removal

ResearchDGX agent

arXiv:2608.04330v1 Announce Type: new Abstract: Removing the left context from a causal language model reveals a useful kind of boundary: an edge where the model processes the same right-hand tokens w

RingSQL: Schema-Independent Synthetic Data Generation for Text-to-SQL Reinforcement Learning

ResearchDGX agent

arXiv:2601.05451v2 Announce Type: replace-cross Abstract: Recent advances in text-to-SQL have been driven by larger models, better datasets, and new training methods like RLVR. However, progress remai

RIP to every 'which framework should I use' thread. harness on top, framework in the middle, runtime at the floor. one stack, three heights,…

Model ReleasesDGX agent

RIP to every 'which framework should I use' thread. harness on top, framework in the middle, runtime at the floor. one stack, three heights, fully composable. paste these four images into claude and a

← Previous
1…8788899091…1408
Next →