AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
22,134 results
Model Releases

InternBootcamp Technical Report: Boosting LLM Reasoning with Verifiable Task Scaling

DGX agent

arXiv:2508.08636v2 Announce Type: replace Abstract: Large language models (LLMs) have revolutionized artificial intelligence by enabling complex reasoning capabilities. While recent advancements in re

model-releasesarxiv-cs-cl
21 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

JobArabi: An Arabic Corpus and Analysis of Job Announcements from Social Media

DGX agent

arXiv:2605.20960v1 Announce Type: new Abstract: This paper introduces JobArabi, a large-scale corpus of Arabic job announcements collected from social media between January 2024 and October 2025. The

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

Lean Refactor: Multi-Objective Controllable Proof Optimization via Agentic Strategy Search

DGX agent

arXiv:2605.20244v1 Announce Type: cross Abstract: We present Lean Refactor, a plug-and-play retrieval-augmented agentic framework for multi-objective, controllable, and version-robust refactoring of L

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

Leveraging Vision-Language Models to Detect Attention in Educational Videos

DGX agent

arXiv:2605.20211v1 Announce Type: new Abstract: Educational videos are a cornerstone of remote and blended learning. However, learners' fluctuating attention remains a significant barrier to effective

model-releasesarxiv-cs-cv
21 May 2026
Applications

Measuring and mitigating overreliance to build human-compatible AI

DGX agent

arXiv:2509.08010v2 Announce Type: replace-cross Abstract: Large language models (LLMs) distinguish themselves from previous technologies by functioning as collaborative ``thought partners,'' capable o

applicationsarxiv-cs-cl
21 May 2026
Model Releases

MemGym: a Long-Horizon Memory Environment for LLM Agents

DGX agent

arXiv:2605.20833v1 Announce Type: new Abstract: Memory is a central capability for LLM agents operating across long-horizon tasks. Existing memory benchmarks predominantly evaluate retention of person

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

MTR-Suite: A Framework for Evaluating and Synthesizing Conversational Retrieval Benchmarks

DGX agent

arXiv:2605.20729v1 Announce Type: new Abstract: Accurate evaluation of conversational retrieval is pivotal for advancing Retrieval-Augmented Generation (RAG) systems. However, existing conversational

model-releasesarxiv-cs-cl
21 May 2026
Applications

Multi-Week, In-Class Deployments of Telepresence Robots With Four Homebound K-12 Students: Benefits, Challenges, and Recommendations

DGX agent

arXiv:2605.20431v1 Announce Type: cross Abstract: Missing significant amounts of school during K-12 education is known to put students' cognitive and social development at risk. Alternatives such as h

applicationsarxiv-cs-ro
21 May 2026
Agents

Retrieval-Augmented Code Generation: A Survey with Focus on Repository-Level Approaches

DGX agent

arXiv:2510.04905v3 Announce Type: replace-cross Abstract: Recent advances in large language models (LLMs) have significantly improved automated code generation. While existing approaches have achieved

agentsarxiv-cs-cl
21 May 2026
Model Releases

ShadeBench: A Benchmark Dataset for Building Shade Simulation in Sustainable Society

DGX agent

arXiv:2605.20510v1 Announce Type: new Abstract: Urban heat exposure is becoming an increasingly critical challenge due to the intensifying urban heat island effect. Fine-grained shade patterns, especi

model-releasesarxiv-cs-cv
21 May 2026
Safety

Shiny Stories, Hidden Struggles: Investigating the Representation of Disability Through the Lens of LLMs

DGX agent

arXiv:2605.20191v1 Announce Type: new Abstract: Modern Large Language Models (LLMs) have recently attracted much attention for their ability to simulate human behavior and generate text that reflects

safetyarxiv-cs-cl
21 May 2026
Agents

Smaller Abstract State Spaces Enable Cross-Scale Generalization in Reinforcement Learning

DGX agent

arXiv:2605.20272v1 Announce Type: new Abstract: While humans readily generalize abstract concepts to more complex or larger tasks, building Reinforcement Learning (RL) systems with this ability remain

agentsarxiv-cs-lg
21 May 2026
Agents

SubTGraph: Large-Scale Subterranean Environment Synthesis with Controllable Topological Variability for Robotic Autonomy Validation

DGX agent

arXiv:2605.20917v1 Announce Type: new Abstract: Subterranean (SubT) environments have been a frontier for autonomous robotics, driven by the push for automation of mining operations and the interest i

agentsarxiv-cs-ro
21 May 2026
Safety

Time-Prompt: Integrated Heterogeneous Prompts for Unlocking LLMs in Time Series Forecasting

DGX agent

arXiv:2506.17631v4 Announce Type: replace Abstract: Time series forecasting aims to model temporal dependencies among variables for future state inference, holding significant importance and widesprea

safetyarxiv-cs-lg
21 May 2026
Model Releases

Towards UAV Detection in the Real World: A New Multispectral Dataset UAVNet-MS and a New Method

DGX agent

arXiv:2605.20963v1 Announce Type: new Abstract: The proliferation of unmanned aerial vehicles (UAVs) has created urgent demand for precise UAV monitoring. Existing RGB-based systems rely on spatial cu

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

Training Language Agents to Learn from Experience

DGX agent

arXiv:2605.20477v1 Announce Type: cross Abstract: Language agents can adapt from experience in interactive environments, but current reflection-based methods can only self-correct within a single task

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

Understanding and Improving Communication Performance in Multi-node LLM Inference

DGX agent

arXiv:2511.09557v4 Announce Type: replace-cross Abstract: As large language models (LLMs) continue to grow in size, distributed inference has become increasingly important. Model-parallel strategies m

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

A Reproducibility Analysis of PO4ISR: Diagnosing and Mitigating Semantic Drift in LLM-Based Session Recommendation

DGX agent

arXiv:2605.18780v1 Announce Type: cross Abstract: Reasoning-based Large Language Models (LLMs) like PO4ISR have set new benchmarks in session-based recommendation. However, the reproducibility of thei

model-releasesarxiv-cs-ai
20 May 2026
Hardware

Accelerating Sparse Transformer Inference on GPU

DGX agent

arXiv:2506.06095v4 Announce Type: replace Abstract: Large language models (LLMs) are popular around the world due to their powerful understanding capabilities. As the core component of LLMs, accelerat

hardwarearxiv-cs-lg
20 May 2026
Tutorials

BCI-sift: An automated feature selection toolbox for Brain Computer Interface applications

DGX agent

arXiv:2605.19646v1 Announce Type: cross Abstract: Advancements in clinical Brain-Computer Interfaces (BCIs) depend on precise and reliable signal interpretation. However, the high-dimensional and nois

tutorialsarxiv-cs-lg
20 May 2026
Model Releases

BLINKG: A Benchmark for LLM-Integrated Knowledge Graph Generation

DGX agent

arXiv:2605.19518v1 Announce Type: new Abstract: Generating Knowledge Graphs (KGs) remains one of the most time-consuming and labor-intensive tasks for knowledge engineers, as they need to identify sem

model-releasesarxiv-cs-ai
20 May 2026
Applications

CAIT: A Syntactic Parsing Toolkit for Child-Adult InTeractions

DGX agent

arXiv:2605.19718v1 Announce Type: new Abstract: CHILDES is a paramount resource for language acquisition studies -- yet computational tools for analyzing its syntactic structure remain limited. Levera

applicationsarxiv-cs-cl
20 May 2026
Model Releases

CogScale: Scalable Benchmark for Sequence Processing

DGX agent

arXiv:2605.19758v1 Announce Type: new Abstract: The ability to maintain and manipulate information over time is a fundamental aspect of living beings and Artificial Intelligence. While modern models h

model-releasesarxiv-cs-ai
20 May 2026
Agents

Decentralized autonomous organization and blockchain-based incentivization framework for community-based facilities management

DGX agent

arXiv:2605.18773v1 Announce Type: cross Abstract: Traditional facility management often relies on centralized decision-making structures that limit stakeholder participation, leading to misalignment w

agentsarxiv-cs-ai
20 May 2026
Agents

Discoverable Agent Knowledge -- A Formal Framework for Agentic KG Affordances (Extended Version)

DGX agent

arXiv:2605.19186v1 Announce Type: new Abstract: Two decades ago, the Semantic Web Services community was asked how agents with different ontological commitments could discover, compose, and invoke web

agentsarxiv-cs-ai
20 May 2026
Safety

Dual-Gated Epistemic Time-Dilation: Autonomous Compute Modulation in Asynchronous MARL

DGX agent

arXiv:2603.23722v2 Announce Type: replace-cross Abstract: While Multi-Agent Reinforcement Learning (MARL) algorithms achieve unprecedented successes across complex continuous domains, their standard d

safetyarxiv-cs-lg
20 May 2026
Model Releases

EgoBabyVLM: Benchmarking Cross-Modal Learning from Naturalistic Egocentric Video Data

DGX agent

arXiv:2605.19130v1 Announce Type: cross Abstract: Children acquire language grounding with remarkable robustness from limited visuo-linguistic input in ways that surpass today's best large multimodal

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Federated Learning for ICD Classification with Lightweight Models and Pretrained Embeddings

DGX agent

arXiv:2507.03122v2 Announce Type: replace-cross Abstract: This study investigates the feasibility and performance of federated learning (FL) for multi-label ICD code classification using clinical note

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

FedMental: Evaluating Federated Learning for Mental Health Detection from Social Media Data

DGX agent

arXiv:2605.18936v1 Announce Type: cross Abstract: Social media text data are often used to train Machine Learning (ML) models to identify users exhibiting high-risk mental health behaviors. However, s

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

Fine-tuning Large Language Model for Automated Algorithm Design

DGX agent

arXiv:2507.10614v2 Announce Type: replace-cross Abstract: The integration of large language models (LLMs) into automated algorithm design has shown promising potential. A prevalent approach embeds LLM

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

FineBench: Benchmarking and Enhancing Vision-Language Models for Fine-grained Human Activity Understanding

DGX agent

arXiv:2605.19846v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) have demonstrated remarkable capabilities in general video understanding, yet they often struggle with the fine-grained

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

HAVEN: Hierarchically Aligned Multimodal Benchmark for Unified Video Understanding

DGX agent

arXiv:2605.19223v1 Announce Type: new Abstract: While Multimodal Large Language Models (MLLMs) exhibit strong performance on standard video tasks, their ability to faithfully summarize and reason over

model-releasesarxiv-cs-cv
20 May 2026
Agents

How to Model AI Agents as Personas?: Applying the Persona Ecosystem Playground to 41,300 Posts on Moltbook for Behavioral Insights

DGX agent

arXiv:2603.03140v3 Announce Type: replace-cross Abstract: AI agents are increasingly active on social media platforms, generating content and interacting with one another at scale. Yet the behavioral

agentsarxiv-cs-ai
20 May 2026
Safety

Improved visual-information-driven model for crowd simulation and its modular application

DGX agent

arXiv:2504.03758v4 Announce Type: replace-cross Abstract: Crowd movement simulation is crucial for pedestrian safety management and facility design. Data-driven models offer the potential to improve r

safetyarxiv-cs-cv
20 May 2026
Model Releases

K-Quantization and its Impact on Output Performance

DGX agent

arXiv:2605.19645v1 Announce Type: new Abstract: Recent advancements in large language models (LLMs) have shown their remarkable capacities in many NLP tasks. However, their substantial size often pres

model-releasesarxiv-cs-cl
20 May 2026
Applications

LLM-MC-Affect: LLM-Based Monte Carlo Modeling of Affective Trajectories and Latent Ambiguity for Interpersonal Dynamic Insight

DGX agent

arXiv:2601.03645v2 Announce Type: replace Abstract: Emotional coordination is a core property of human interaction that shapes how relational meaning is constructed in real time. While text-based affe

applicationsarxiv-cs-cl
20 May 2026
Model Releases

MSAVBench: Towards Comprehensive and Reliable Evaluation of Multi-Shot Audio-Video Generation

DGX agent

arXiv:2605.20183v1 Announce Type: new Abstract: Video generation is rapidly evolving from single-shot synthesis to complex multi-shot audio-video (MSAV) narratives to meet real-world demands. However,

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

NGL: Natural Garment Language for Training-Free Sewing Pattern Estimation

DGX agent

arXiv:2602.20700v2 Announce Type: replace Abstract: Estimating sewing patterns from images is a practical approach for creating high-quality 3D garments, but it remains challenging due to the scarcity

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

Physics-in-the-Loop: A Hybrid Agentic Architecture for Validated CAD Engineering Design

DGX agent

arXiv:2605.19717v1 Announce Type: new Abstract: Large Language Models (LLMs) can generate Computer-Aided Design (CAD), yet lack physical comprehension required for reliable engineering design. Instead

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

PlantTraitNet: An Uncertainty-Aware Multimodal Framework for Global-Scale Plant Trait Inference from Citizen Science Data

DGX agent

arXiv:2511.06943v3 Announce Type: replace-cross Abstract: Global plant maps of plant traits, such as leaf nitrogen or plant height, are essential for understanding ecosystem processes, including the c

model-releasesarxiv-cs-ai
20 May 2026
Applications

Position: Graph Condensation Needs a Reset -- Move Beyond Full-dataset Training and Model-Dependence

DGX agent

arXiv:2605.18893v1 Announce Type: new Abstract: Graph Neural Networks (GNNs) are powerful tools for learning from graph-structured data, but their scalability is increasingly strained by the size of r

applicationsarxiv-cs-lg
20 May 2026
Safety

Position: Let's Develop Data Probes to Fundamentally Understand How Data Affects LLM Performance

DGX agent

arXiv:2605.18801v1 Announce Type: new Abstract: Data is fundamental to large language models (LLMs). However, understanding of what makes certain data useful for different stages of an LLM workflow, i

safetyarxiv-cs-ai
20 May 2026
Safety

Position: Uncertainty Quantification in LLMs is Just Unsupervised Clustering

DGX agent

arXiv:2605.19220v1 Announce Type: cross Abstract: Uncertainty Quantification (UQ) is widely regarded as the primary safeguard for deploying Large Language Models (LLMs) in high-stakes domains. However

safetyarxiv-cs-ai
20 May 2026
Safety

Precision Physical Activity Prescription via Reinforcement Learning for Functional Actions

DGX agent

arXiv:2605.19208v1 Announce Type: cross Abstract: Physical activity (PA) plays an important role in maintaining and improving health. Daily steps have been a key PA measure that is easily accessible w

safetyarxiv-cs-lg
20 May 2026
Model Releases

PromptRad: Knowledge-Enhanced Multi-Label Prompt-Tuning for Low-Resource Radiology Report Labeling

DGX agent

arXiv:2605.20052v1 Announce Type: cross Abstract: Automatic report labeling facilitates the identification of clinical findings from unstructured text and enables large-scale annotation for medical im

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Quantifying the Generalization Gap in Seizure Detection: A Large-Scale Empirical Benchmark via the SzCORE Challenge

DGX agent

arXiv:2505.18191v2 Announce Type: replace-cross Abstract: Reliable automatic seizure detection from long-term electroencephalography (EEG) remains an unsolved challenge, as current models often fail t

model-releasesarxiv-cs-ai
20 May 2026
Safety

Rapid patient-specific neural networks for intraoperative X-ray to volume registration

DGX agent

arXiv:2503.16309v2 Announce Type: replace-cross Abstract: Advanced navigation techniques in image-guided interventions and surgical robotics require the rapid and precise alignment of 3D preoperative

safetyarxiv-cs-cv
20 May 2026
Model Releases

Rewriting History: A Recipe for Interventional Analyses to Study Data Effects on Model Behavior

DGX agent

arXiv:2510.14261v2 Announce Type: replace Abstract: We present an experimental recipe for studying the relationship between training data and language model (LM) behavior. We outline steps for interve

model-releasesarxiv-cs-cl
20 May 2026
← Previous
1…436437438439440…462
Next →