AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,457
  • Agents7,399
  • Applications5,302
  • Concepts5
  • Hardware1,786
  • Industry6,117
  • Local Ai4,835
  • Model Releases23,193
  • Research19,715
  • Safety13,094
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,457
  • Agents7,399
  • Applications5,302
  • Concepts5
  • Hardware1,786
  • Industry6,117
  • Local Ai4,835
  • Model Releases23,193
  • Research19,715
  • Safety13,094
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
Human
86,457Total entries
1Added by human
86,456Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
61,498 results
3 Jul 2026

On the Role of Computation in Reinforcement Learning

SafetyDGX agent

arXiv:2602.05999v4 Announce Type: replace Abstract: How does the amount of compute available to a reinforcement learning (RL) policy affect its learning? Can policies using a fixed amount of parameter

On the Role of Directionality in Structural Generalization

ResearchDGX agent

arXiv:2607.02307v1 Announce Type: new Abstract: Several SLOG test categories explicitly involve directional distinctions (modifier position shifts, argument extraction positions), yet AM-Parser, the p

On the Sample Efficiency of Inverse Dynamics Models for Semi-Supervised Imitation Learning

SafetyDGX agent

arXiv:2602.02762v2 Announce Type: replace Abstract: Semi-supervised imitation learning (SSIL) consists in learning a policy from a small dataset of action-labeled trajectories and a much larger datase

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

On the Utility and Factual Reliability of Pruned Mixture-of-Experts Models in the Biomedical Domain

Model ReleasesDGX agent

arXiv:2607.01444v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) models offer inference speedups via selective activation but impose substantial memory requirements because the whole network

One Demonstration Is Enough for Real-World Robotic Reinforcement Learning

SafetyDGX agent

arXiv:2607.01651v1 Announce Type: new Abstract: Learning effective robot control policies on physical hardware is challenging due to costly data collection and the difficulty of reward specification.

One More Time: Revisiting Neural Quantum States from a Reinforcement Learning Perspective

Model ReleasesDGX agent

arXiv:2607.02292v1 Announce Type: new Abstract: Neural quantum states (NQS) provide a flexible and scalable framework for approximating quantum many-body wavefunctions. Among NQS parameterizations, au

Online Resource Allocation with Continuous Random Consumption: Regret under Degeneracy

SafetyDGX agent

arXiv:2607.02196v1 Announce Type: new Abstract: We study online resource allocation when both rewards and consumption sizes may be continuously distributed. Requests arrive sequentially and must be ac

Online Safety Monitoring for LLMs

SafetyDGX agent

arXiv:2607.02510v1 Announce Type: new Abstract: Despite alignment training, LLMs remain prone to generating unsafe outputs at deployment time. Monitoring outputs online and raising an alarm when safet

OntoLearner: A Modular Python Library for Ontology Learning with Large Language Models

ResearchDGX agent

arXiv:2607.01977v1 Announce Type: new Abstract: Ontology learning (OL) aims to automatically construct structured knowledge models from text, yet progress remains fragmented across methods, domains, a

OpenSafeIntent: Evaluating Intent-Calibrated Safe Completion Across Dual-Use Prompt Sets

Model ReleasesDGX agent

arXiv:2607.02047v1 Announce Type: cross Abstract: Safe completion requires models to provide useful assistance without enabling harm, but this behavior is difficult to evaluate with isolated prompts.

Ophiuchus: Incentivizing Tool-augmented 'Think with Images' for Joint Medical Segmentation, Understanding and Reasoning

AgentsDGX agent

arXiv:2512.14157v2 Announce Type: replace Abstract: Recent medical MLLMs have made significant progress in generating step-by-step textual reasoning chains. However, they still struggle with complex c

OPINE-World: Programmatic World Modeling with Ontology-error-Prioritized Interactive Exploration

Model ReleasesDGX agent

arXiv:2607.01531v1 Announce Type: new Abstract: Learning how an environment behaves from interaction is central to building agents that adapt to unfamiliar tasks. World models learned with deep networ

Optimal Stabilizer Testing and Learning with Limited Quantum Memory

TutorialsDGX agent

arXiv:2607.02444v1 Announce Type: cross Abstract: We study stabilizer state testing and learning with limited coherent quantum memory. Here an algorithm sequentially receives copies of an unknown n-qu

Optimizing RAG Rerankers with LLM Feedback via Reinforcement Learning

ResearchDGX agent

arXiv:2604.02091v2 Announce Type: replace-cross Abstract: Rerankers play a pivotal role in refining retrieval results for Retrieval-Augmented Generation. However, current reranking models are typicall

Optimizing Visual Generative Models via Distribution-wise Rewards

SafetyDGX agent

arXiv:2607.02291v1 Announce Type: new Abstract: Conventional reinforcement learning strategies for visual generation typically employ sample-wise reward functions, yet this practice frequently results

OrbitQuant: Data-Agnostic Quantization for Image and Video Diffusion Transformers

Model ReleasesDGX agent

arXiv:2607.02461v1 Announce Type: cross Abstract: Diffusion transformers (DiTs) achieve state-of-the-art image and video generation, but their multi-step sampling and growing parameter count make infe

Overthink-Triggered Slowdown Attacks on LVLM-Based Robotic Systems

SafetyDGX agent

arXiv:2607.01518v1 Announce Type: cross Abstract: Large Vision-Language Models (LVLMs) have been increasingly integrated into robotic systems. However, these models may exhibit overthinking behaviors,

Overview of Risk Assessment and Management for Intelligent Systems under the AI Act and Beyond

ResearchDGX agent

arXiv:2607.02197v1 Announce Type: cross Abstract: The society and emerging risk-based regulatory frameworks for AI underscore the need for rigorous risk assessment to ensure safe and reliable AI syste

PACE: A Neuro-Symbolic Framework for Plausible and Actionable Counterfactual Explanations

ApplicationsDGX agent

arXiv:2607.01306v1 Announce Type: new Abstract: Counterfactual explanations explain machine learning predictions by identifying minimal input changes that would alter a model's decision. Although many

PACE: A Proxy for Agentic Capability Evaluation

Model ReleasesDGX agent

arXiv:2607.02032v1 Announce Type: new Abstract: Evaluating LLM agents on benchmarks like SWE-Bench and GAIA can be expensive, time-consuming, and requires complex infrastructure. A single evaluation c

PairCoder++: Pair Programming as a Universal Paradigm for Verified Code-Driven Multimodal and Structured-Artifact Generation

Model ReleasesDGX agent

arXiv:2607.01883v1 Announce Type: new Abstract: Code is the medium through which large language models generate structured artifacts: charts, scientific figures, vector graphics, CAD models, 3D scenes

Parameter Golf: What Really Works?

Model ReleasesDGX agent

arXiv:2607.01517v1 Announce Type: new Abstract: How far can a language model improve under a strict artifact budget? Parameter Golf posed this question as an open community challenge in which particip

PARTREP: Learning What to Repeat for Decoder-only LLMs

ResearchDGX agent

arXiv:2607.01792v1 Announce Type: new Abstract: While decoder-only LLMs excel at a vast array of natural language tasks, it suffers from an asymmetric information flow induced by causal attention: lat

Path-level Hindsight Instructions for Semantic Exploration in Vision-Language Navigation

Model ReleasesDGX agent

arXiv:2607.01754v1 Announce Type: new Abstract: On-policy exploration is a crucial component for training robust Vision-Language Navigation agents, as it exposes the policy to a broader state distribu

Path planning for unmanned naval surface vehicles

ResearchDGX agent

arXiv:2607.01631v1 Announce Type: new Abstract: There nowadays is a myriad of approaches to real-time avoidance of fixed obstacles for unmanned surface vehicles (USVs) and, to a lesser extent, also th

Phonikud: Overcoming Phonetic Underspecification for Hebrew Text-To-Speech

Model ReleasesDGX agent

arXiv:2506.12311v4 Announce Type: replace Abstract: Text-to-speech (TTS) for Modern Hebrew is challenged by the language's orthographic complexity, with existing solutions ignoring underspecified phon

PhysMani: Physics-principled 3D World Model for Dynamic Object Manipulation

Model ReleasesDGX agent

arXiv:2607.01938v1 Announce Type: cross Abstract: Manipulating fast and dynamically moving targets in unstructured 3D environments remains challenging for embodied AI. Existing visual-language-action

PixGS: Pixel-Space Diffusion for Direct 3D Gaussian Splat Generation

HardwareDGX agent

arXiv:2607.01803v1 Announce Type: cross Abstract: Recent advances in 3D content generation from text or images have achieved impressive results, yet view inconsistency from 2D generators and the scarc

Playing 20 Question Game with Policy-Based Reinforcement Learning

SafetyDGX agent

arXiv:1808.07645v5 Announce Type: replace-cross Abstract: The 20 Questions (Q20) game is a well known game which encourages deductive reasoning and creativity. In the game, the answerer first thinks o

Pmeta-TLA: Backdoor Attacks for Speech Classification Models via Meta-Learning with Timbre Leakage Attack

ResearchDGX agent

arXiv:2607.01702v1 Announce Type: cross Abstract: Recently, speech classification methods have gained widespread adoption in intelligent gadgets. Current study indicates that backdoor attacks provide

Population-Based Multi-Objective Training of Discriminators for Semi-Supervised GANs

ResearchDGX agent

arXiv:2607.01907v1 Announce Type: cross Abstract: Semi-supervised generative adversarial networks (SSL-GANs) can exploit large unlabeled datasets while retaining a classifier in the discriminator, but

Population-Scale Segmentation of Penile Tissue in DIXON MRI using Deep Learning for Quantitative Phenotyping in Male Reproductive Health

Model ReleasesDGX agent

arXiv:2607.02127v1 Announce Type: cross Abstract: Penile measurement is clinically relevant across male reproductive and urogenital health, including conditions such as micropenis, congenital and endo

Power Systems Agent Benchmark: Executable Evaluation of AI Agents in Electric Power Engineering

Model ReleasesDGX agent

arXiv:2606.20950v2 Announce Type: replace Abstract: Executable evaluation -- checking the consequences of an agent's actions with a program rather than grading its prose -- has become a prominent way

PPTArena: A Benchmark for PowerPoint Editing

Model ReleasesDGX agent

arXiv:2512.03042v3 Announce Type: replace-cross Abstract: We introduce PPTArena, a benchmark for PowerPoint editing that evaluates how agents modify real slides from natural-language instructions. Unl

Pre-Flight: A Benchmark for Evaluating Large Language Models on Aviation Operational Knowledge

Model ReleasesDGX agent

arXiv:2607.01829v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly proposed for aviation business operations, from documentation and training generation to customer facing a

Predicting Closed-Loop Performance of Latent World Models: Offline Checkpoint Selection for MPC and Model-Based RL Under Non-Markovian Rewards in LunarLander

SafetyDGX agent

arXiv:2607.01736v1 Announce Type: cross Abstract: We study how to predict the downstream closed-loop performance of a learned latent world model from validation-time diagnostics alone. Choosing the ri

Predicting Early Stages Of Alzheimer's Disease And Identifying Key Biomarkers Using Deep Artificial Neural Network And Ensemble Of Machine Learning Methodologies

ResearchDGX agent

arXiv:2607.02142v1 Announce Type: cross Abstract: Alzheimers disease (AD) is a brain disorder that develops slowly and mainly affects memory, thinking, language, and daily activities. It is one of the

Prediction Sets for Counterfactual Decisions: Coverage, Optimality, and Conformal Prediction

SafetyDGX agent

arXiv:2607.02206v1 Announce Type: cross Abstract: Predictions are increasingly used to guide high-stakes decisions, from treatment selection to policy making. To ensure reliability with imperfect pred

Predictive Conformal Slip Monitoring: An Empirical Evaluation of Rolling Split Conformal Prediction for Pre-Incident Traction Loss Detection

ResearchDGX agent

arXiv:2607.02124v1 Announce Type: new Abstract: Conventional traction control architectures intervene only after the adhesion limit of a tire has already been breached. This paper investigates whether

PreScience: A Dataset and Benchmark for Scientific Forecasting

Model ReleasesDGX agent

arXiv:2602.20459v2 Announce Type: replace Abstract: Can AI systems trained on the existing scientific record forecast the advances that will follow? We introduce PreScience, a dataset and benchmark fo

Privacy-Preserving and Verifiable Approximate Distributed Coded Computing

ResearchDGX agent

arXiv:2607.02187v1 Announce Type: new Abstract: Distributed machine learning enables collaborative model training without centralizing data, but it also exposes learning processes to privacy leakage a

Probabilistic Low-Voltage Peak Load Forecasting with Time Series Foundation Models Evaluated on Application-Oriented Metrics

ApplicationsDGX agent

arXiv:2607.01966v1 Announce Type: new Abstract: Low-voltage load forecasting is an important component in current and future energy systems with a high degree of electrification and decentralized gene

Probing Chemical Language Models: Effects of Pre-training and Fine-tuning

ResearchDGX agent

arXiv:2607.02140v1 Announce Type: new Abstract: Chemical language models (CLMs) are trained with linearized representations such as SMILES, yet it remains unclear which chemically meaningful substruct

Probing Spectrum-Like Organization of States of Mind in Transformer Representation Spaces

ResearchDGX agent

arXiv:2512.22227v3 Announce Type: replace Abstract: We investigate whether graded states of mind form spectrum-like structure in transformer representation spaces. To do so, we construct a dataset of

ProCal: Inference-Time Proposal Calibration for Open-Vocabulary Object Detection

Local AiDGX agent

arXiv:2607.01759v1 Announce Type: cross Abstract: Open-vocabulary object detection aims to localize and classify objects beyond the fixed set of categories seen dur ing training. Recent open-vocabular

Procedural Memory Distillation: Online Reflection for Self-Improving Language Models

Local AiDGX agent

arXiv:2607.01480v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR), along with recent selfdistillation variants such as SDPO, evaluates each rollout against a verifi

Profit-Based Counterfactual Explanations for Product Improvement: A Case Study of Manga Sales in Japan

ApplicationsDGX agent

arXiv:2607.01610v1 Announce Type: new Abstract: Counterfactual explanation (CE) is widely used to enhance the interpretability of machine learning models and support data-driven decision-making based

Program-as-Weights: A Programming Paradigm for Fuzzy Functions

Model ReleasesDGX agent

arXiv:2607.02512v1 Announce Type: cross Abstract: Many everyday programming tasks resist clean rule-based implementation, such as alerting on important log lines, repairing malformed JSON, or ranking

Prompt Coverage Adequacy

AgentsDGX agent

arXiv:2607.02057v1 Announce Type: cross Abstract: In recent years, it has become increasingly evident that large language models (LLMs) and autonomous agents raise the level of abstraction in software

Prompt Framing Distorts Count-Based Evaluation of LLM Error Detection: Evidence from Numeric Anchoring

Model ReleasesDGX agent

arXiv:2607.01240v1 Announce Type: cross Abstract: Count-based F1 is widely used as a proxy for LLM error-detection quality, but this paper shows that it can rise dramatically without a corresponding i

Provably Finding a Hidden Dense Submatrix among Many Planted Dense Submatrices via Convex Programming

ApplicationsDGX agent

arXiv:2601.03946v3 Announce Type: replace-cross Abstract: We consider the densest submatrix problem, which seeks the submatrix of fixed size of a given binary matrix that contains the most nonzero ent

ProWAFT: A ROMA-LPD Instance for Workload-Aware and Dynamic Fault Tolerance in FPGA-Based CNN Accelerators

ResearchDGX agent

arXiv:2607.01602v1 Announce Type: new Abstract: SRAM-based FPGAs provide an attractive platform for energy- and latency-constrained CNN inference at the network edge, yet transient faults can lead to

Psychological Imagination Networks Show Cross-Population Centrality and Clustering Alignment in Humans That Large Language Models Fail to Replicate

Model ReleasesDGX agent

arXiv:2510.04391v5 Announce Type: replace Abstract: Mental imagery vividness is a stable individual trait, yet whether imagined scenarios share relational structure across human and synthetic large la

Psychological Steering in LLMs: An Evaluation of Effectiveness and Trustworthiness

Model ReleasesDGX agent

arXiv:2510.04484v2 Announce Type: replace-cross Abstract: The ability to control LLMs' emulated emotional states and personality traits is an essential step in enabling rich, human-centered interactio

Purified OPSD: On-Policy Self-Distillation Without Losing How to Think

SafetyDGX agent

arXiv:2607.02234v1 Announce Type: new Abstract: On-policy self-distillation (OPSD) has emerged as a promising paradigm for improving LLM reasoning, where a privileged teacher with access to reference

Q-GAIN: A Python Package for Machine Learning and Physically Informed Analysis Applications

ResearchDGX agent

arXiv:2607.02413v1 Announce Type: cross Abstract: Here we describe the quantum gas analysis and inference (Q-GAIN) Python package, which enables rapid deployment of machine learning (ML) and physics-i

QFedAgent: Quantum-Enhanced Personalized Federated Learning for Multi-Agent Activity Recognition

Model ReleasesDGX agent

arXiv:2607.02426v1 Announce Type: cross Abstract: Federated learning (FL) enables collaborative model training across distributed devices without sharing raw data, making it suitable for privacy-sensi

QTALE: Quantization-Robust Token-Adaptive Layer Execution for LLMs

ResearchDGX agent

arXiv:2602.10431v4 Announce Type: replace Abstract: Large language models (LLMs) demand substantial computational and memory resources, posing challenges for efficient deployment. Two complementary ap

QuadRocket: An Aerial Robotic Testbed for Adaptive Thrust-Vector Control of Rocket-Like Vehicles

ResearchDGX agent

arXiv:2607.02474v1 Announce Type: new Abstract: This paper presents QuadRocket, a quadrotor-based rocket prototype that provides a low-cost, low-risk platform for validating advanced thrust-vector con

Quantifying the Uncertainty of Blindly Estimated Room Embeddings Using a Dispersion-Calibrated Score

SafetyDGX agent

arXiv:2607.01527v1 Announce Type: cross Abstract: Room embeddings derived from reverberant speech are often unreliable: speech content and recording degradation can alter the representation even when

← Previous
1…282283284285286…1025
Next →