AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
Human
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
59,295 results
16 Apr 2026

Estimating Continuous Treatment Effects with Two-Stage Kernel Ridge Regression

SafetyDGX agent

arXiv:2604.13410v1 Announce Type: cross Abstract: We study the problem of estimating the effect function for a continuous treatment, which maps each treatment value to a population-averaged outcome. A

Evaluating LLM-Based Translation of a Low-Resource Technical Language: The Medical and Philosophical Greek of Galen

Model ReleasesDGX agent

arXiv:2602.24119v2 Announce Type: replace Abstract: Purpose: This study evaluates the quality of commercial large language model (LLM) machine translation (MT) for Ancient Greek technical prose and be

Evaluating Supervised Machine Learning Models: Principles, Pitfalls, and Metric Selection

Model ReleasesDGX agent
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.13882v1 Announce Type: new Abstract: The evaluation of supervised machine learning models is a critical stage in the development of reliable predictive systems. Despite the widespread avail

Evaluating the Evaluator: Problems with SemEval-2020 Task 1 for Lexical Semantic Change Detection

Model ReleasesDGX agent

arXiv:2604.13232v1 Announce Type: new Abstract: This discussion paper re-examines SemEval-2020 Task 1, the most influential shared benchmark for lexical semantic change detection, through a three-part

Evaluating the Formal Reasoning Capabilities of Large Language Models through Chomsky Hierarchy

Model ReleasesDGX agent

arXiv:2604.02709v2 Announce Type: replace Abstract: The formal reasoning capabilities of LLMs are crucial for advancing automated software engineering. However, existing benchmarks for LLMs lack syste

EVE: A Domain-Specific LLM Framework for Earth Intelligence

Model ReleasesDGX agent

arXiv:2604.13071v1 Announce Type: new Abstract: We introduce Earth Virtual Expert (EVE), the first open-source, end-to-end initiative for developing and deploying domain-specialized LLMs for Earth Int

Event-Adaptive State Transition and Gated Fusion for RGB-Event Object Tracking

ResearchDGX agent

arXiv:2604.13426v1 Announce Type: new Abstract: Existing Vision Mamba-based RGB-Event(RGBE) tracking methods suffer from using static state transition matrices, which fail to adapt to variations in ev

Event Tensor: A Unified Abstraction for Compiling Dynamic Megakernel

HardwareDGX agent

arXiv:2604.13327v1 Announce Type: cross Abstract: Modern GPU workloads, especially large language model (LLM) inference, suffer from kernel launch overheads and coarse synchronization that limit inter

Evolvable Embodied Agent for Robotic Manipulation via Long Short-Term Reflection and Optimization

SafetyDGX agent

arXiv:2604.13533v1 Announce Type: cross Abstract: Achieving general-purpose robotics requires empowering robots to adapt and evolve based on their environment and feedback. Traditional methods face li

Explainable Fall Detection for Elderly Care via Temporally Stable SHAP in Skeleton-Based Human Activity Recognition

Local AiDGX agent

arXiv:2604.13279v1 Announce Type: new Abstract: Fall detection in elderly care requires not only accurate classification but also reliable explanations that clinicians can trust. However, existing pos

Exploring Urban Land Use Patterns by Pattern Mining and Unsupervised Learning

ResearchDGX agent

arXiv:2604.13050v1 Announce Type: cross Abstract: Urban areas are intricate systems shaped by socioeconomic, environmental, and infrastructural factors, with land use patterns serving as aspects of ur

Exposia: Teaching and Assessment of Academic Writing Skills for Research Project Proposals and Peer Feedback

Model ReleasesDGX agent

arXiv:2601.06536v2 Announce Type: replace Abstract: We present Exposia, the first public dataset that connects writing and feedback in higher education, enabling research on educationally grounded com

ExpSeek: Self-Triggered Experience Seeking for Web Agents

Model ReleasesDGX agent

arXiv:2601.08605v2 Announce Type: replace Abstract: Experience intervention in web agents emerges as a promising technical paradigm, enhancing agent interaction capabilities by providing valuable insi

F-Actor: Controllable Conversational Behaviour in Full-Duplex Models

Model ReleasesDGX agent

arXiv:2601.11329v3 Announce Type: replace Abstract: Spoken conversational systems require more than accurate speech generation to have human-like conversations: to feel natural and engaging, they must

Failure Identification in Imitation Learning Via Statistical and Semantic Filtering

SafetyDGX agent

arXiv:2604.13788v1 Announce Type: cross Abstract: Imitation learning (IL) policies in robotics deliver strong performance in controlled settings but remain brittle in real-world deployments: rare even

Failure Makes the Agent Stronger: Enhancing Accuracy through Structured Reflection for Reliable Tool Interactions

Model ReleasesDGX agent

arXiv:2509.18847v3 Announce Type: replace-cross Abstract: Tool-augmented large language models (LLMs) are usually trained with supervised imitation or coarse-grained reinforcement learning that optimi

FAST: A Synergistic Framework of Attention and State-space Models for Spatiotemporal Traffic Prediction

ResearchDGX agent

arXiv:2604.13453v1 Announce Type: new Abstract: Traffic forecasting requires modeling complex temporal dynamics and long-range spatial dependencies over large sensor networks. Existing methods typical

Fast and Simple Densest Subgraph with Predictions

ApplicationsDGX agent

arXiv:2505.12600v3 Announce Type: replace-cross Abstract: We study the densest subgraph problem and its NP-hard densest at-most-k subgraph variant through the lens of learning-augmented algorithms. We

Fast training of accurate physics-informed neural networks without gradient descent

Model ReleasesDGX agent

arXiv:2405.20836v3 Announce Type: replace-cross Abstract: Solving time-dependent Partial Differential Equations (PDEs) is one of the most critical problems in computational science. While Physics-Info

Fast Voxelization and Level of Detail for Microgeometry Rendering

Local AiDGX agent

arXiv:2604.13191v1 Announce Type: cross Abstract: Many materials show anisotropic light scattering patterns due to the shape and local alignment of their underlying micro structures: surfaces with sma

FCBV-Net: Category-Level Robotic Garment Smoothing via Feature-Conditioned Bimanual Value Prediction

SafetyDGX agent

arXiv:2508.05153v2 Announce Type: replace Abstract: Category-level generalization for robotic garment manipulation, such as bimanual smoothing, remains a significant hurdle due to high dimensionality,

Feed-Forward 3D Scene Modeling: A Problem-Driven Perspective

ApplicationsDGX agent

arXiv:2604.14025v1 Announce Type: new Abstract: Reconstructing 3D representations from 2D inputs is a fundamental task in computer vision and graphics, serving as a cornerstone for understanding and i

Feedback-Driven Tool-Use Improvements in Large Language Models via Automated Build Environments

Local AiDGX agent

arXiv:2508.08791v3 Announce Type: replace Abstract: Effective tool use is essential for large language models (LLMs) to interact with their environment. However, progress is limited by the lack of eff

FieldWorkArena: Agentic AI Benchmark for Real Field Work Tasks

Model ReleasesDGX agent

arXiv:2505.19662v3 Announce Type: replace-cross Abstract: This paper introduces FieldWorkArena, a benchmark for agentic AI targeting real-world field work. With the recent increase in demand for agent

FiLM-Nav: Efficient and Generalizable Navigation via VLM Fine-tuning

Model ReleasesDGX agent

arXiv:2509.16445v2 Announce Type: replace Abstract: Enabling robotic assistants to navigate complex environments and locate objects described in free-form language is a critical capability for real-wo

First-See-Then-Design: A Multi-Stakeholder View for Optimal Performance-Fairness Trade-Offs

SafetyDGX agent

arXiv:2604.14035v1 Announce Type: new Abstract: Fairness in algorithmic decision-making is often defined in the predictive space, where predictive performance - used as a proxy for decision-maker (DM)

FlexGuard: Continuous Risk Scoring for Strictness-Adaptive LLM Content Moderation

Model ReleasesDGX agent

arXiv:2602.23636v3 Announce Type: replace Abstract: Ensuring the safety of LLM-generated content is essential for real-world deployment. Most existing guardrail models formulate moderation as a fixed

Flow-based Generative Modeling of Potential Outcomes and Counterfactuals

Model ReleasesDGX agent

arXiv:2505.16051v4 Announce Type: replace-cross Abstract: Predicting potential and counterfactual outcomes from observational data is central to individualized decision-making, particularly in clinica

Fluids You Can Trust: Property-Preserving Operator Learning for Incompressible Flows

HardwareDGX agent

arXiv:2602.15472v4 Announce Type: replace-cross Abstract: We present a novel property-preserving kernel-based operator learning method for incompressible flows governed by the incompressible Navier--S

fMRI-LM: Towards a Universal Foundation Model for Language-Aligned fMRI Understanding

Model ReleasesDGX agent

arXiv:2511.21760v3 Announce Type: replace Abstract: Recent advances in multimodal large language models (LLMs) have enabled unified reasoning across images, audio, and video, but extending such capabi

Foresight Optimization for Strategic Reasoning in Large Language Models

SafetyDGX agent

arXiv:2604.13592v1 Announce Type: new Abstract: Reasoning capabilities in large language models (LLMs) have generally advanced significantly. However, it is still challenging for existing reasoning-ba

Form Without Function: Agent Social Behavior in the Moltbook Network

AgentsDGX agent

arXiv:2604.13052v1 Announce Type: cross Abstract: Moltbook is a social network where every participant is an AI agent. We analyze 1,312,238 posts, 6.7~million comments, and over 120,000 agent profiles

Free Geometry: Refining 3D Reconstruction from Longer Versions of Itself

Model ReleasesDGX agent

arXiv:2604.14048v1 Announce Type: new Abstract: Feed-forward 3D reconstruction models are efficient but rigid: once trained, they perform inference in a zero-shot manner and cannot adapt to the test s

Free Lunch for Unified Multimodal Models: Enhancing Generation via Reflective Rectification with Inherent Understanding

ResearchDGX agent

arXiv:2604.13540v1 Announce Type: new Abstract: Unified Multimodal Models (UMMs) aim to integrate visual understanding and generation within a single structure. However, these models exhibit a notable

From Alignment to Prediction: A Study of Self-Supervised Learning and Predictive Representation Learning

SafetyDGX agent

arXiv:2604.13518v1 Announce Type: new Abstract: Self-supervised learning has emerged as a major technique for the task of learning from unlabeled data, where the current methods mostly revolve around

From Anchors to Supervision: Memory-Graph Guided Corpus-Free Unlearning for Large Language Models

Local AiDGX agent

arXiv:2604.13777v1 Announce Type: new Abstract: Large language models (LLMs) may memorize sensitive or copyrighted content, raising significant privacy and legal concerns. While machine unlearning has

From Feelings to Metrics: Understanding and Formalizing How Users Vibe-Test LLMs

Model ReleasesDGX agent

arXiv:2604.14137v1 Announce Type: new Abstract: Evaluating LLMs is challenging, as benchmark scores often fail to capture models' real-world usefulness. Instead, users often rely on ``vibe-testing'':

From Instruction to Event: Sound-Triggered Mobile Manipulation

SafetyDGX agent

arXiv:2601.21667v2 Announce Type: replace-cross Abstract: Current mobile manipulation research predominantly follows an instruction-driven paradigm, where agents rely on predefined textual commands to

From Order to Distribution: A Spectral Characterization of Forgetting in Continual Learning

ResearchDGX agent

arXiv:2604.13460v1 Announce Type: new Abstract: A central challenge in continual learning is forgetting, the loss of performance on previously learned tasks induced by sequential adaptation to new one

From Pixels to Nucleotides: End-to-End Token-Based Video Compression for DNA Storage

ResearchDGX agent

arXiv:2604.13667v1 Announce Type: new Abstract: DNA-based storage has emerged as a promising approach to the global data crisis, offering molecular-scale density and millennial-scale stability at low

From Plausibility to Verifiability: Risk-Controlled Generative OCR with Vision-Language Models

Model ReleasesDGX agent

arXiv:2603.19790v3 Announce Type: replace Abstract: Modern vision-language models (VLMs) can act as generative OCR engines, yet open-ended decoding can expose rare but consequential failures. We ident

From Prediction to Justification: Aligning Sentiment Reasoning with Human Rationale via Reinforcement Learning

ResearchDGX agent

arXiv:2604.13398v1 Announce Type: new Abstract: While Aspect-based Sentiment Analysis (ABSA) systems have achieved high accuracy in identifying sentiment polarities, they often operate as 'black boxes

From P(y|x) to P(y): Investigating Reinforcement Learning in Pre-train Space

SafetyDGX agent

arXiv:2604.14142v1 Announce Type: cross Abstract: While reinforcement learning with verifiable rewards (RLVR) significantly enhances LLM reasoning by optimizing the conditional distribution P(y|x), it

From Relevance to Authority: Authority-aware Generative Retrieval in Web Search Engines

ApplicationsDGX agent

arXiv:2604.13468v1 Announce Type: cross Abstract: Generative information retrieval (GenIR) formulates the retrieval process as a text-to-text generation task, leveraging the vast knowledge of large la

From Seeing it to Experiencing it: Interactive Evaluation of Intersectional Voice Bias in Human-AI Speech Interaction

SafetyDGX agent

arXiv:2604.13067v1 Announce Type: cross Abstract: SpeechLLMs process spoken language directly from audio, but accent and vocal identity cues can lead to biased behaviour. Current bias evaluations ofte

From Synchrony to Sequence: Exo-to-Ego Generation via Interpolation

ResearchDGX agent

arXiv:2604.13793v1 Announce Type: new Abstract: Exo-to-Ego video generation aims to synthesize a first-person video from a synchronized third-person view and corresponding camera poses. While paired s

From Weights to Activations: Is Steering the Next Frontier of Adaptation?

Model ReleasesDGX agent

arXiv:2604.14090v1 Announce Type: new Abstract: Post-training adaptation of language models is commonly achieved through parameter updates or input-based methods such as fine-tuning, parameter-efficie

From Where Words Come: Efficient Regularization of Code Tokenizers Through Source Attribution

SafetyDGX agent

arXiv:2604.14053v1 Announce Type: new Abstract: Efficiency and safety of Large Language Models (LLMs), among other factors, rely on the quality of tokenization. A good tokenizer not only improves infe

Frozen Forecasting: A Unified Evaluation

ResearchDGX agent

arXiv:2507.13942v2 Announce Type: replace Abstract: Forecasting future events is a fundamental capability for general-purpose systems that plan or act across different levels of abstraction. Yet, eval

Functional Emotions or Situational Contexts? A Discriminating Test from the Mythos Preview System Card

Model ReleasesDGX agent

arXiv:2604.13466v1 Announce Type: cross Abstract: The Claude Mythos Preview system card deploys emotion vectors, sparse autoencoder (SAE) features, and activation verbalisers to study model internals

Gaslight, Gatekeep, V1-V3: Early Visual Cortex Alignment Shields Vision-Language Models from Sycophantic Manipulation

Model ReleasesDGX agent

arXiv:2604.13803v1 Announce Type: new Abstract: Vision-language models are increasingly deployed in high-stakes settings, yet their susceptibility to sycophantic manipulation remains poorly understood

Geminet: Learning the Duality-based Iterative Process for Lightweight Traffic Engineering in Changing Topologies

ResearchDGX agent

arXiv:2506.23640v2 Announce Type: replace-cross Abstract: Recently, researchers have explored ML-based Traffic Engineering (TE), leveraging neural networks to solve TE problems traditionally addressed

Generalization Guarantees on Data-Driven Tuning of Gradient Descent with Langevin Updates

TutorialsDGX agent

arXiv:2604.13130v1 Announce Type: new Abstract: We study learning to learn for regression problems through the lens of hyperparameter tuning. We propose the Langevin Gradient Descent Algorithm (LGD),

GeoBridge: A Semantic-Anchored Multi-View Foundation Model Bridging Images and Text for Geo-Localization

Model ReleasesDGX agent

arXiv:2512.02697v3 Announce Type: replace Abstract: Cross-view geo-localization infers a location by retrieving geo-tagged reference images that visually correspond to a query image. However, the trad

GeoLink: A 3D-Aware Framework Towards Better Generalization in Cross-View Geo-Localization

Local AiDGX agent

arXiv:2604.13183v1 Announce Type: new Abstract: Generalizable cross-view geo-localization aims to match the same location across views in unseen regions and conditions without GPS supervision. Its cor

Geometric Context Transformer for Streaming 3D Reconstruction

ResearchDGX agent

arXiv:2604.14141v1 Announce Type: new Abstract: Streaming 3D reconstruction aims to recover 3D information, such as camera poses and point clouds, from a video stream, which necessitates geometric acc

GeoVision-Enabled Digital Twin for Hybrid Autonomous-Teleoperated Medical Responses

AgentsDGX agent

arXiv:2604.13248v1 Announce Type: new Abstract: Remote medical response systems are increasingly being deployed to support emergency care in disaster-affected and infrastructure-limited environments.

Getting the Numbers Rightnicode{x2014}Modelling Multi-Class Object Counting in Dense and Varied Scenes

ResearchDGX agent

arXiv:2510.02213v2 Announce Type: replace Abstract: Density map estimation enables accurate object counting in heavily occluded, and densely packed scenes where detection-based counting fails. In mult

Giving Voice to the Constitution: Low-Resource Text-to-Speech for Quechua and Spanish Using a Bilingual Legal Corpus

ApplicationsDGX agent

arXiv:2604.13288v1 Announce Type: new Abstract: We present a unified pipeline for synthesizing high-quality Quechua and Spanish speech for the Peruvian Constitution using three state-of-the-art text-t

Goal2Skill: Long-Horizon Manipulation with Adaptive Planning and Reflection

AgentsDGX agent

arXiv:2604.13942v1 Announce Type: new Abstract: Recent vision-language-action (VLA) systems have demonstrated strong capabilities in embodied manipulation. However, most existing VLA policies rely on

← Previous
1…917918919920921…989
Next →