AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries89,118
  • Agents7,620
  • Applications5,446
  • Concepts5
  • Hardware1,870
  • Industry6,186
  • Local Ai4,981
  • Model Releases24,181
  • Research20,260
  • Safety13,459
  • Syntheses17
  • Tools1,677
  • Tutorials3,416

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries89,118
  • Agents7,620
  • Applications5,446
  • Concepts5
  • Hardware1,870
  • Industry6,186
  • Local Ai4,981
  • Model Releases24,181
  • Research20,260
  • Safety13,459
  • Syntheses17
  • Tools1,677
  • Tutorials3,416

Source
Human
89,118Total entries
1Added by human
89,117Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
89,117 results
26 May 2026

Benchmarking non-conformity score functions in conformal prediction

ResearchDGX agent

arXiv:2605.24983v1 Announce Type: new Abstract: Conformal prediction is a useful and versatile alternative to model calibration in machine learning classification. It replaces single-class prediction

Benchmarking Patent Embeddings: A Multi-Task Evaluation of 22 Models Across Retrieval, Classification, and Clustering

Model ReleasesDGX agent

arXiv:2605.24297v1 Announce Type: cross Abstract: Which fine-tuning signals improve patent embedding models, and do gains transfer across patent landscapes? We benchmark 22 embedding models, from 22M-

Benchmarking Pathology Foundation Models for Spatial Domain Understanding

Model ReleasesDGX agent
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.25764v1 Announce Type: cross Abstract: Pathology foundation models (PFMs) have emerged as a core approach for learning transferable representations from whole slide images (WSIs), and they

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork

Model ReleasesDGX agent

arXiv:2605.24423v1 Announce Type: new Abstract: In-Context Reinforcement Learning (ICRL) has enabled foundation agents to adapt instantaneously to novel tasks, yet its efficacy in Ad-Hoc Teamwork (AHT

Better, Faster: Harnessing Self-Improvement in Large Reasoning Models

ResearchDGX agent

arXiv:2605.24998v1 Announce Type: new Abstract: Self-improvement training enables the large reasoning models (LRMs) to improve themselves by self-generating reasoning trajectories as training data wit

Beyond Control-Flow: Integrating the Resource Perspective into Multi-Collaborative Process Modeling from Text

ResearchDGX agent

arXiv:2605.24546v1 Announce Type: new Abstract: Process modeling is a sub-domain of Business Process Management (BPM) focused on the translation of process artifacts into formal models. This task trad

Beyond Final Answers: Auditing Trajectory-Level Hallucinations in Multi-Agent Industrial Workflows

Model ReleasesDGX agent

arXiv:2605.24219v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly deployed as autonomous agents that reason, use tools, and act over multiple steps. Yet most hallucination

Beyond Fixed Points: Superpolynomial Capacity of Asymmetric Hopfield Networks

ResearchDGX agent

arXiv:2605.24611v1 Announce Type: new Abstract: Classical Hopfield networks are limited to static patterns due to symmetric weights, whereas asymmetric networks can encode temporal sequences via limit

Beyond Generative Priors: Minority Sampling with JEPA-Guided Diffusion

ApplicationsDGX agent

arXiv:2605.24631v1 Announce Type: cross Abstract: Minority sampling aims to generate low-density instances on a data manifold and is of central importance in applications such as medical diagnosis, an

Beyond Inference-Only Deployment: Comparing Weight-Based Consolidation Against Cascading Compaction

Model ReleasesDGX agent

arXiv:2605.24657v1 Announce Type: new Abstract: Major LLM platforms deploy models in an inference-only configuration: the model serves requests but never updates per-user weights. Users must repeatedl

Beyond Killer Robots: General AI Attitudes and Public Support for Military AI in Nine Countries

SafetyDGX agent

arXiv:2605.25196v1 Announce Type: cross Abstract: AI-enabled military systems are a fixture of modern military conflict. Applications vary from autonomous drones for surveillance and attack to AI-supp

Beyond Literal Translation: Evaluating Cultural Effectiveness in Social Media UGC

Model ReleasesDGX agent

arXiv:2605.25626v1 Announce Type: new Abstract: Social media platforms enable large-scale cross-lingual communication, but translating user-generated content (UGC) remains challenging due to its infor

Beyond Predefined Learning Objects: A Thinking-Learning Interaction Model for Up-to-Date Autonomous Robot Learning

AgentsDGX agent

arXiv:2605.23987v1 Announce Type: new Abstract: Autonomous robots operating in open and changing environments cannot always rely on predefined inputs, outputs, and action routines. Although existing l

Beyond Query Memorization: Large Language Model Routing with Query Decomposition and Historical Matching

Model ReleasesDGX agent

arXiv:2605.25558v1 Announce Type: new Abstract: Optimizing the trade-off among predictive performance and computational cost is a central focus in the deployment of Large Language Models (LLMs). Curre

Beyond Static Uncertainty: Modeling Temporal Uncertainty Dynamics for Probabilistic Time Series Forecasting

ApplicationsDGX agent

arXiv:2603.24254v2 Announce Type: replace-cross Abstract: Real-world time series exhibit temporally structured uncertainty: volatility clusters in turbulent regimes, dissipates in stable periods, and

Beyond Summaries: Structure-Aware Labeling of Code Changes with Large Language Models

Model ReleasesDGX agent

arXiv:2605.26100v1 Announce Type: cross Abstract: Code review is a critical practice in software engineering, yet the growing scale and frequency of code patches in modern projects, together with the

Beyond the Aggregation Dilemma: Prior-Retaining Decoupled Learning for Multimodal Graphs

HardwareDGX agent

arXiv:2605.24684v1 Announce Type: cross Abstract: Multimodal Attributed Graph Learning (MAGL) integrates intrinsic node attributes with structural topology via graph aggregation. However, as pretraine

Beyond the Frontier: Stochastic Backtracking for Efficient Test-Time Scaling

ResearchDGX agent

arXiv:2605.25143v1 Announce Type: new Abstract: Test-time scaling improves language model reasoning by spending additional compute to explore multiple solution trajectories. The key challenge is to ma

Beyond the Proxy: Trajectory-Distilled Guidance for Offline GFlowNet Training

SafetyDGX agent

arXiv:2505.20110v3 Announce Type: replace-cross Abstract: Generative Flow Networks (GFlowNets) excel at sampling diverse, high-reward objects. In many practical applications where active reward querie

Beyond the Target: From Imitation to Collaboration in Speculative Decoding

SafetyDGX agent

arXiv:2605.24793v1 Announce Type: new Abstract: Speculative decoding (SPD) accelerates large language model (LLM) inference by letting a smaller draft model propose multiple future tokens that are ver

BigMac: Breaking the Pareto Frontier of Compute and Memory in Multimodal LLM Training

ResearchDGX agent

arXiv:2605.25451v1 Announce Type: new Abstract: Training multimodal large language models (MLLMs) is challenged by both model and data heterogeneity. Existing systems redesign the training pipeline to

Bilevel Optimization of Synthetic Trajectories for Multi-Turn LLM Fine-Tuning

ResearchDGX agent

arXiv:2605.24743v1 Announce Type: cross Abstract: While LLMs excel at single-turn generation, they struggle with long-horizon, multi-turn interactions. Offline reinforcement learning (RL) offers a sca

Binding Visual Features Point by Point

ResearchDGX agent

arXiv:2605.25427v1 Announce Type: cross Abstract: Despite success on standard benchmarks, vision language models display persistent failures on tasks involving processing of multi-object scenes, inclu

BlitzRank: Principled Zero-shot Ranking Agents with Tournament Graphs

ApplicationsDGX agent

arXiv:2602.05448v4 Announce Type: replace Abstract: Selecting the top m from n items via expensive k-wise comparisons is central to settings ranging from LLM-based document reranking to crowdsourced e

Blocked Gibbs meets Diffusion Transformers: Unsupervised Learning for Constraint Optimization

SafetyDGX agent

arXiv:2605.25129v1 Announce Type: new Abstract: Diffusion models have shown promise in learning to solve constraint optimization problems. However, they are mostly restricted to problems with binary v

Board Mode is here. Instead of one long agent session, you now get a creative workspace where research, slides, websites, apps, and design d…

AgentsDGX agent

Board Mode is here. Instead of one long agent session, you now get a creative workspace where research, slides, websites, apps, and design drafts can live side by side. Try it now: https://agent.ii.in

BODHI: Precise OS Kernel Specification Inference

Model ReleasesDGX agent

arXiv:2605.23931v1 Announce Type: new Abstract: The formal verification of operating system kernels requires precise specifications that capture the intended behavior of system calls. Writing these sp

Boosting Inference with Guided Reasoning: Stochastic Exploration for Recursive Models

Local AiDGX agent

arXiv:2605.25230v1 Announce Type: new Abstract: Recent work on recursive architectures has shown that tiny neural networks can be surprisingly powerful on structured reasoning tasks. The trick is to m

Box barely beats the Street’s expectations and nervy investors back away

IndustryDGX agent

Shares of the cloud content management firm Box Inc. headed lower in extended trading today after it barely scraped past Wall Street’s earnings and revenue targets in its first quarter. The company re

BoxLitE: A Faithful Knowledge Base Embedding Based on Convex Optimization

TutorialsDGX agent

arXiv:2605.23937v1 Announce Type: new Abstract: Knowledge base (KB) embeddings aim at combining the capability of classical knowledge graph embeddings to generalize the information present in facts, t

Branch Scaling Manifests as Implicit Architectural Regularization for Improving Generalization in Overparameterized ResNets

ApplicationsDGX agent

arXiv:2403.04545v3 Announce Type: replace Abstract: Scaling factors in residual branches have emerged as a prevalent method for boosting neural network performance, especially in normalization-free ar

Branched Signature Kernel Solvers for ODEs with rough Single-Trajectory signals

ApplicationsDGX agent

arXiv:2605.25826v1 Announce Type: cross Abstract: We develop a branched signature kernel solver for linear and nonlinear ordinary differential equations driven by a single observed trajectory of a pos

Breaking the Chains of Probability: Neutrosophic Logic as a New Framework for Epistemic Uncertainty in Large Language Models

ResearchDGX agent

arXiv:2605.24053v1 Announce Type: new Abstract: Large Language Models (LLMs) are predominantly governed by probabilistic frameworks in which the sum of outcome probabilities is constrained to unity. T

Bridging Earth and Space: A Survey on HAPS for Non-Terrestrial Networks

AgentsDGX agent

arXiv:2510.19731v3 Announce Type: replace-cross Abstract: HAPS are emerging as key enablers in the evolution of 6G wireless networks, bridging terrestrial and non-terrestrial infrastructures. Operatin

Bridging Evolutionary Algorithms and Reinforcement Learning: A Comprehensive Survey on Hybrid Algorithms

ResearchDGX agent

arXiv:2401.11963v5 Announce Type: replace-cross Abstract: Evolutionary Reinforcement Learning (ERL), which integrates Evolutionary Algorithms (EAs) and Reinforcement Learning (RL) for optimization, ha

Bridging On-Device and Cloud LLMs for Collaborative Reasoning: A Unified Methodology for Local Routing and Post-Training

Model ReleasesDGX agent

arXiv:2509.24050v4 Announce Type: replace Abstract: Device-cloud collaboration holds promise for deploying large language models (LLMs), leveraging lightweight on-device models for efficiency while re

Bridging the Gap: Enabling Soft Actor Critic for High Performance Legged Locomotion

SafetyDGX agent

arXiv:2605.24975v1 Announce Type: cross Abstract: Proximal Policy Optimization (PPO) has become the de facto standard for training legged robots, thanks to its robustness and scalability in massively

Bridging the Semantic-Action Gap in Visual Token Pruning for Efficient VLA Inference

ResearchDGX agent

arXiv:2511.16449v4 Announce Type: replace-cross Abstract: Vision-Language-Action (VLA) models have shown great potential for embodied AI by integrating visual perception, language understanding, and a

btw this will be the 3 year anniversary of the Rise of the AI Engineer blogpost. the industry keeps growing and growing, kinda scary to real…

ToolsDGX agent

btw this will be the 3 year anniversary of the Rise of the AI Engineer blogpost. the industry keeps growing and growing, kinda scary to realize i may never top it. AIE EU, AIE MIA and AIE SG reached 8

Build an enterprise observability solution for Amazon Quick

ApplicationsDGX agent

When hundreds to thousands of users are onboarded to an enterprise AI platform, business leaders and platform owners need visibility into who is using the platform, whether users are satisfied with th

Build high-performance generative AI systems with Strands Agents, NVIDIA NIM, and Amazon Bedrock AgentCore

HardwareDGX agent

In this post you'll learn how to build a multi-agent campaign review system that demonstrates parallel reasoning, context persistence, and traceable execution paths using an integrated architecture th

Build highly scalable serverless LangGraph multi-agent systems in AWS with Amazon Bedrock AgentCore

AgentsDGX agent

In this post, we provide a solution to build highly scalable, serverless multi-agent generative AI systems on AWS using LangGraph Agents as orchestrators integrated with Amazon Bedrock AgentCore Memor

Building agents (like software) is a deeply iterative process --> this is why we try to supply as much easy to use tooling and infra as poss…

AgentsDGX agent

Building agents (like software) is a deeply iterative process --> this is why we try to supply as much easy to use tooling and infra as possible so builders can start on their agent improvement journe

Building an Adversarial Malware Dataset by Family and Type: Generation, Evasion, and Poisoning Evaluation

Model ReleasesDGX agent

arXiv:2605.25937v1 Announce Type: cross Abstract: We present a dataset of adversarial malware samples derived from the public RawMal-TF collection of real-world malware binaries. Using a suite of adve

By Their Fruits You Will Know Them: Comparing Formalizations of Law by the Decisions They Encode

ApplicationsDGX agent

arXiv:2605.25186v1 Announce Type: cross Abstract: Formalizing legal provisions promises machine-accessible law and automated legal reasoning, and recent LLMs make it tempting to generate such formaliz

Byzantine-Robust Federated Learning with Learnable Aggregation Weights

ResearchDGX agent

arXiv:2511.03529v2 Announce Type: replace Abstract: Federated Learning (FL) enables clients to collaboratively train a global model without sharing their private data. However, the presence of malicio

CAFD: Concept-Aware DNN Fault Detection using VLMs

ApplicationsDGX agent

arXiv:2605.24008v1 Announce Type: new Abstract: Fault detection for Deep Neural Networks (DNNs) has received increasing attention in recent years. While more advanced hybrid approaches have been propo

CAffNet: Hard Constraint-Affine Neural Networks

ResearchDGX agent

arXiv:2605.24437v1 Announce Type: new Abstract: We present a novel framework for embedding hard constraint satisfaction into neural network (NN) architectures, specifically feedforward neural networks

CALIBURN: A Regime-Sensitivity Study of Operationally Calibrated Streaming Intrusion Detection

ApplicationsDGX agent

arXiv:2605.24696v1 Announce Type: cross Abstract: Streaming network intrusion detection systems must process flows continuously while keeping memory bounded, but most current methods leave alerting th

Can Large Language Models Resolve Semantic Discrepancy in Self-Destructive Subcultures? Evidence from Jirai Kei

SafetyDGX agent

arXiv:2601.05004v2 Announce Type: replace Abstract: Self-destructive behaviors are linked to complex psychological states and can be challenging to diagnose. These behaviors may be even harder to iden

Can LLMs Time Travel? Enhancing Temporal Consistency in Legal Agentic Search through Reinforcement Learning

Model ReleasesDGX agent

arXiv:2605.25920v1 Announce Type: cross Abstract: While large language models (LLMs) augmented with agentic search capabilities show promise for legal reasoning, they overlook a fundamental constraint

Can LoRA Fusion Support Cross-Domain Tasks in Cloud-Edge Collaboration?

Model ReleasesDGX agent

arXiv:2605.23913v1 Announce Type: cross Abstract: Cloud-hosted large language models (LLMs) commonly rely on LoRA for domain adaptation, yet domain data are distributed across multiple edge devices an

Capability and Robustness Cannot Both Be Free: An Information-Theoretic Bound for Vision-Language-Action Models

SafetyDGX agent

arXiv:2605.25889v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models are increasingly deployed on real robots, where each predicted action is executed and each failure carries a safet

Capture-Calibrate-Coach: A Graph-Based Framework for Knowledge Monitoring Estimation and Adaptive Feedback

ResearchDGX agent

arXiv:2605.25419v1 Announce Type: new Abstract: Effective learning support requires understanding not only what learners know but also how accurately they perceive their own understanding. This metaco

CARL-CXR: Continual Adapter-Based Routing for Task-Unknown Chest Radiograph Classification

ResearchDGX agent

arXiv:2602.15811v2 Announce Type: replace-cross Abstract: Clinical deployment of chest radiograph classifiers requires models that can be updated as new datasets become available without retraining on

Cascade-KDE: Robust Time-Series Restoration under Out-of-Distribution Impulse Corruptions

Model ReleasesDGX agent

arXiv:2605.24055v1 Announce Type: cross Abstract: Real-world time-series data in industrial sensing, healthcare, and energy systems is often corrupted by a mixture of Gaussian noise and occasional lar

Catching MRI outliers: unsupervised detection and localization of MRI artefacts and clinical anomalies using deep learning

Local AiDGX agent

arXiv:2605.24609v1 Announce Type: cross Abstract: Artificial intelligence is increasingly integrated into radiotherapy workflows, yet such pipelines remain vulnerable to out-of-distribution image data

Catching The Correct Answer Trap: Characterising AI Tutor Blind Spots When Analysing Student Reasoning

ResearchDGX agent

arXiv:2605.23925v1 Announce Type: cross Abstract: Intelligent tutoring systems increasingly provide automated feedback on student work, but robust feedback requires assessing reasoning, not only final

Causal methods for LLM development and evaluation

SafetyDGX agent

arXiv:2605.25998v1 Announce Type: new Abstract: Large language model (LLM) development is currently driven by large-scale empirical iteration over data mixtures, reward models, routing strategies, and

Causal Tongue-Tie: LLMs Can Encode Causal Direction, But Their Yes/No Outputs Fail to Express

Model ReleasesDGX agent

arXiv:2605.25891v1 Announce Type: cross Abstract: We find a mismatch between what large language models encode about a causal question and what they answer. On anti-commonsense CLadder items, a fixed

← Previous
1…820821822823824…1486
Next →