AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “tools”

GridTimelineEvolution
9,955 results
22 Apr 2026

Workspace agents

TutorialsDGX agent

Workspace agents are AI systems designed to automate tasks and workflows within collaborative work environments, likely covering how these agents can handle scheduling, communication, document managem

21 Apr 2026

A Discordance-Aware Multimodal Framework with Multi-Agent Clinical Reasoning

AgentsDGX agent

arXiv:2604.16333v1 Announce Type: new Abstract: Knee osteoarthritis frequently exhibits discordance between structural damage observed in imaging and patient-reported symptoms such as pain. This misma

A Two-Stage Deep Learning Framework for Segmentation of Ten Gastrointestinal Organs from Coronal MR Enterography

ResearchDGX agent

arXiv:2604.17118v1 Announce Type: cross Abstract: Accurate segmentation of gastrointestinal (GI) organs in magnetic resonance enterography (MRE) is critical for diagnosing inflammatory bowel disease (

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Agentic Risk-Aware Set-Based Engineering Design

Model ReleasesDGX agent

arXiv:2604.16687v1 Announce Type: cross Abstract: This paper introduces a multi-agent framework guided by Large Language Models (LLMs) to assist in the early stages of engineering design, a phase ofte

Agents Explore but Agents Ignore: LLMs Lack Environmental Curiosity

AgentsDGX agent

arXiv:2604.17609v1 Announce Type: new Abstract: LLM-based agents are assumed to integrate environmental observations into their reasoning: discovering highly relevant but unexpected information should

AI Data Transformation Guide for Data Engineers and Data Scientists

TutorialsDGX agent

This guide from Databricks covers data transformation techniques and best practices essential for preparing data for AI/ML projects, addressing workflows that both data engineers and data scientists e

Alignment Data Map for Efficient Preference Data Selection and Diagnosis

SafetyDGX agent

arXiv:2505.23114v3 Announce Type: replace Abstract: Human preference data is essential for aligning large language models (LLMs) with human values, but collecting such data is often costly and ineffic

Anyone here using ai agent orchestration software to control multiple hermes agents? I'm retired and have some extra hardware

AgentsDGX agent

This Reddit post from r/ollama asks the community about AI agent orchestration software for managing multiple Hermes agents, posted by a retired individual with available hardware resources. The post

Aspect Ratios & Resolution in ChatGPT Images 2.0, demonstrated by @dibyayB

Model ReleasesDGX agent

ChatGPT Images 2.0 supports multiple aspect ratios and resolutions for image generation, allowing users greater flexibility in creating images tailored to different use cases and display formats. The

b8864

Local AiDGX agent

b8864 is a build release of llama.cpp, an open-source C/C++ library for large language model inference. The project uses rapid release cycles with frequent build tags published as intermediate develop

BenchMarker: An Education-Inspired Toolkit for Highlighting Flaws in Multiple-Choice Benchmarks

Model ReleasesDGX agent

arXiv:2602.06221v2 Announce Type: replace Abstract: Multiple-choice question answering (MCQA) is standard in NLP, but benchmarks lack rigorous quality control. We present BenchMarker, an education-ins

Beyond 'I Don't Know': Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty

Model ReleasesDGX agent

arXiv:2604.17293v1 Announce Type: new Abstract: Reliable Large Language Models (LLMs) should abstain when confidence is insufficient. However, prior studies often treat refusal as a generic 'I don't k

Beyond Static Benchmarks: Synthesizing Harmful Content via Persona-based Simulation for Robust Evaluation

ResearchDGX agent

arXiv:2604.17020v1 Announce Type: new Abstract: Static benchmarks for harmful content detection face limitations in scalability and diversity, and may also be affected by contamination from web-scale

BhashaSutra: A Task-Centric Unified Survey of Indian NLP Datasets, Corpora, and Resources

ResearchDGX agent

arXiv:2604.18423v1 Announce Type: new Abstract: India's linguistic landscape, spanning 22 scheduled languages and hundreds of marginalized dialects, has driven rapid growth in NLP datasets, benchmarks

BIASEDTALES-ML: A Multilingual Dataset for Analyzing Narrative Attribute Distributions in LLM-Generated Stories

SafetyDGX agent

arXiv:2604.17008v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly used to generate narrative content, including children's stories, which play an important role in social a

Bridging the Culture Gap: A Framework for LLM-Driven Socio-Cultural Localization of Math Word Problems in Low-Resource Languages

Local AiDGX agent

arXiv:2508.14913v4 Announce Type: replace Abstract: Large language models (LLMs) have demonstrated significant capabilities in solving mathematical problems expressed in natural language. However, mul

Building the foundation for AI across the public sector through our partner ecosystem

HardwareDGX agent

The demand for AI within the public sector has never been higher. Practitioners and CXO’s are looking for ways to harness AI to improve mission outcomes, enhance security, and streamline operations. H

Can we generate portable representations for clinical time series data using LLMs?

ApplicationsDGX agent

arXiv:2603.23987v2 Announce Type: replace Abstract: Deploying clinical ML is slow and brittle: models that work at one hospital often degrade under distribution shifts at the next. In this work, we st

Chaos-Enhanced Prototypical Networks for Few-Shot Medical Image Classification

ResearchDGX agent

arXiv:2604.17300v1 Announce Type: cross Abstract: The scarcity of labeled clinical data in oncology makes Few-Shot Learning (FSL) a critical framework for Computer Aided Diagnostics, but we observed t

ClawEnvKit: Automatic Environment Generation for Claw-Like Agents

Model ReleasesDGX agent

arXiv:2604.18543v1 Announce Type: cross Abstract: Constructing environments for training and evaluating claw-like agents remains a manual, human-intensive process that does not scale. We argue that wh

CoDial: Interpretable Task-Oriented Dialogue Systems Through Dialogue Flow Alignment

Model ReleasesDGX agent

arXiv:2506.02264v3 Announce Type: replace Abstract: Building Task-Oriented Dialogue (TOD) systems that generalize across different tasks remains a challenging problem. Data-driven approaches often str

COSEARCH: Joint Training of Reasoning and Document Ranking via Reinforcement Learning for Agentic Search

SafetyDGX agent

arXiv:2604.17555v1 Announce Type: cross Abstract: Agentic search -- the task of training agents that iteratively reason, issue queries, and synthesize retrieved information to answer complex questions

DEM Refinement and Validation on the Lunar Surface Using Shape-from-Shading with Chandrayaan-2 OHRC Imagery

Model ReleasesDGX agent

arXiv:2604.17436v1 Announce Type: new Abstract: This study presents a Shape from Shading (SfS) framework to enhance sub-metre resolution lunar digital elevation models (DEMs) using imagery from the Or

DreamShot: Personalized Storyboard Synthesis with Video Diffusion Prior

SafetyDGX agent

arXiv:2604.17195v1 Announce Type: new Abstract: Storyboard synthesis plays a crucial role in visual storytelling, aiming to generate coherent shot sequences that visually narrate cinematic events with

EmoVerse: A MLLMs-Driven Emotion Representation Dataset for Interpretable Visual Emotion Analysis

ResearchDGX agent

arXiv:2511.12554v2 Announce Type: replace Abstract: Visual Emotion Analysis (VEA) aims to bridge the affective gap between visual content and human emotional responses. Despite its promise, progress i

Enabling Stroke-Level Structural Analysis of Hieroglyphic Scripts without Language-Specific Priors

ResearchDGX agent

arXiv:2601.05508v2 Announce Type: replace-cross Abstract: Hieroglyphs, as logographic writing systems, encode rich semantic and cultural information within their internal structural composition. Yet,

End-to-End Optimization of LLM-Driven Multi-Agent Search Systems via Heterogeneous-Group-Based Reinforcement Learning

SafetyDGX agent

arXiv:2506.02718v2 Announce Type: replace Abstract: Large language models (LLMs) are versatile, yet their deployment in complex real-world settings is limited by static knowledge cutoffs and the diffi

Enhancing Trust in Large Language Models via Uncertainty-Calibrated Fine-Tuning

ResearchDGX agent

arXiv:2412.02904v2 Announce Type: replace Abstract: Large language models (LLMs) have revolutionized the field of natural language processing with their impressive reasoning and question-answering cap

ESsEN: Training Compact Discriminative Vision-Language Transformers in a Low-Resource Setting

Model ReleasesDGX agent

arXiv:2604.18452v1 Announce Type: cross Abstract: Vision-language modeling is rapidly increasing in popularity with an ever expanding list of available models. In most cases, these vision-language mod

FLiP: Towards understanding and interpreting multimodal multilingual sentence embeddings

Model ReleasesDGX agent

arXiv:2604.18109v1 Announce Type: new Abstract: This paper presents factorized linear projection (FLiP) models for understanding pretrained sentence embedding spaces. We train FLiP models to recover t

From Fallback to Frontline: When Can LLMs be Superior Annotators of Human Perspectives?

ResearchDGX agent

arXiv:2604.17968v1 Announce Type: cross Abstract: Although large language models (LLMs) are increasingly used as annotators at scale, they are typically treated as a pragmatic fallback rather than a f

FSEVAL: Feature Selection Evaluation Toolbox and Dashboard

ResearchDGX agent

arXiv:2604.18227v1 Announce Type: new Abstract: Feature selection is a fundamental machine learning and data mining task, involved with discriminating redundant features from informative ones. It is a

GEAR up to get the most out of AI learning at Google Cloud Next ‘26

AgentsDGX agent

Our new GEAR program, powered by Google Skills, equips every professional with the hands-on AI training needed to build and launch enterprise-ready agents at scale. Anyone can join GEAR for access to

Given the significant progress over the past year, this year’s festival will likely mark a tipping point. Submissions are still open!

IndustryDGX agent

Given the significant progress over the past year, this year’s festival will likely mark a tipping point. Submissions are still open! The Runway AI Festival returns this June to NY and LA to celebrate

Graph neural network for colliding particles with an application to sea ice floe modeling

TutorialsDGX agent

arXiv:2602.16213v2 Announce Type: replace-cross Abstract: This paper introduces a novel approach to sea ice modeling using Graph Neural Networks (GNNs), utilizing the natural graph structure of sea ic

HopWeaver: Cross-Document Synthesis of High-Quality and Authentic Multi-Hop Questions

ResearchDGX agent

arXiv:2505.15087v3 Announce Type: replace Abstract: Multi-Hop Question Answering (MHQA) is crucial for evaluating the model's capability to integrate information from diverse sources. However, creatin

How Much Data is Enough? The Zeta Law of Discoverability in Biomedical Data, featuring the enigmatic Riemann zeta function

ResearchDGX agent

arXiv:2604.17581v1 Announce Type: new Abstract: How much data is enough to make a scientific discovery? As biomedical datasets scale to millions of samples and AI models grow in capacity, progress inc

Hyperspectral Unmixing Hierarchies

ResearchDGX agent

arXiv:2604.16969v1 Announce Type: new Abstract: Unmixing reveals the spatial distribution and spectral details of different constituents, called endmembers, in a hyperspectral image. Because unmixing

I hope this helps. If you need more help let me know: Connecting OpenClaw to the X API is straightforward now thanks to X’s official native …

Model ReleasesDGX agent

I hope this helps. If you need more help let me know: Connecting OpenClaw to the X API is straightforward now thanks to X’s official native support... The best and most direct method uses the official

I think the CoALA paper's classification system of semantic/episodic/procedural is maybe the closest thing we have to a standard for agent m…

AgentsDGX agent

The CoALA paper proposes a classification system for agent memory that distinguishes between semantic memory (facts and concepts), episodic memory (specific experiences and events), and procedural mem

IYKYK (But AI Doesn't): Automated Content Moderation Does Not Capture Communities' Heterogeneous Attitudes Towards Reclaimed Language

SafetyDGX agent

arXiv:2604.16654v1 Announce Type: new Abstract: Reclaimed slur usage is a common and meaningful practice online for many marginalized communities. It serves as a source of solidarity, identity, and sh

Judge a Book by its Cover: Investigating Multi-Modal LLMs for Multi-Page Handwritten Document Transcription

Model ReleasesDGX agent

arXiv:2502.20295v2 Announce Type: replace-cross Abstract: Handwriting text recognition (HTR) remains a challenging task. Existing approaches require fine-tuning on labeled data, which is impractical t

📢 Kimi K2.6 API is live • Input Price (Cache Hit): 0.16 / M tokens • Input Price (Cache Miss): 0.95 / M tokens • Output: $4.00 / M tokens…

Model ReleasesDGX agent

📢 Kimi K2.6 API is live • Input Price (Cache Hit): 0.16 / M tokens • Input Price (Cache Miss): 0.95 / M tokens • Output: $4.00 / M tokens Kimi K2.6 is our latest + most intelligent model - stronger lo

Kimi K2.6 autonomously overhauled exchange-core, an 8-year-old open-source financial matching engine. Over a 13-hour execution, the model it…

Model ReleasesDGX agent

Kimi K2.6 autonomously overhauled exchange-core, an 8-year-old open-source financial matching engine. Over a 13-hour execution, the model iterated through 12 optimization strategies, initiating over 1

Kimi K2.6 demonstrates strong long-horizon coding in complex engineering tasks: Kimi K2.6 successfully downloaded and deployed the Qwen3.5-0…

Local AiDGX agent

Kimi K2.6 demonstrates strong long-horizon coding in complex engineering tasks: Kimi K2.6 successfully downloaded and deployed the Qwen3.5-0.8B model locally on a Mac. By implementing and optimizing m

Kimi K2.6 is now live inside Anything!

Model ReleasesDGX agent

Kimi K2.6, an AI model from Moonshot, has been integrated into the Anything platform. This update likely enables users to access Kimi's capabilities directly within the Anything application interface.

Large Language Models Are Bad Dice Players: LLMs Struggle to Generate Random Numbers from Statistical Distributions

ApplicationsDGX agent

arXiv:2601.05414v2 Announce Type: replace Abstract: As large language models (LLMs) transition from chat interfaces to integral components of stochastic pipelines and systems approaching general intel

Learning to Seek Help: Dynamic Collaboration Between Small and Large Language Models

Local AiDGX agent

arXiv:2604.17827v1 Announce Type: new Abstract: Large language models (LLMs) offer strong capabilities but raise cost and privacy concerns, whereas small language models (SLMs) facilitate efficient an

Linking Exteroception and Proprioception through Improved Contact Modeling for Soft Growing Robots

ResearchDGX agent

arXiv:2507.10694v2 Announce Type: replace Abstract: Passive deformation due to compliance is a commonly used benefit of soft robots, providing opportunities to achieve robust actuation with few active

Live LTL Progress Tracking: Towards Task-Based Exploration

AgentsDGX agent

arXiv:2604.17106v1 Announce Type: new Abstract: Motivated by the challenge presented by non-Markovian objectives in reinforcement learning (RL), we present a novel framework to track and represent the

Love this work from Aksel and the post-training team at Hugging Face! Turns out the HF ecosystem (papers, datasets, models all accessible th…

Model ReleasesDGX agent

Love this work from Aksel and the post-training team at Hugging Face! Turns out the HF ecosystem (papers, datasets, models all accessible through CLI, skills and md files) is perfect for running SOTA

LVLMs and Humans Ground Differently in Referential Communication

ResearchDGX agent

arXiv:2601.19792v3 Announce Type: replace Abstract: For generative AI agents to partner effectively with human users, the ability to accurately predict human intent is critical. But this ability to co

Matrix: Peer-to-Peer Multi-Agent Synthetic Data Generation Framework

AgentsDGX agent

arXiv:2511.21686v2 Announce Type: replace Abstract: Synthetic data has become increasingly important for training large language models, especially when real data is scarce, expensive, or privacy-sens

Muscle-inspired magnetic actuators that push, pull, crawl, and grasp

HardwareDGX agent

arXiv:2604.18090v1 Announce Type: new Abstract: Functional magnetic composites capable of large deformation, load bearing, and multifunctional motion are essential for next-generation adaptive soft ro

need open standards

AgentsDGX agent

The post likely discusses the importance of open standards in AI and software development, advocating for transparency and interoperability rather than proprietary solutions. Harrison Chase, co-founde

Nothing better than sprinting on a day-0 release alongside partners that run just as fast. Thank you @sarahmsachs and the @NotionHQ crew. Ex…

AgentsDGX agent

Nothing better than sprinting on a day-0 release alongside partners that run just as fast. Thank you @sarahmsachs and the @NotionHQ crew. Excited to see the day 1 reactions continue to roll in today.

On the Convergence and Size Transferability of Continuous-depth Graph Neural Networks

SafetyDGX agent

arXiv:2510.03923v2 Announce Type: replace Abstract: Continuous-depth graph neural networks, also known as Graph Neural Differential Equations (GNDEs), combine the structural inductive bias of Graph Ne

On the Predictive Power of Representation Dispersion in Language Models

Model ReleasesDGX agent

arXiv:2506.24106v2 Announce Type: replace Abstract: We show that a language model's ability to predict text is tightly linked to the breadth of its embedding space: models that spread their contextual

PAC-Bayes Bounds for Gibbs Posteriors via Singular Learning Theory

Model ReleasesDGX agent

arXiv:2604.17219v1 Announce Type: cross Abstract: We derive explicit non-asymptotic PAC-Bayes generalization bounds for Gibbs posteriors, that is, data-dependent distributions over model parameters ob

Pearmut: Human Evaluation of Translation Made Trivial

ResearchDGX agent

arXiv:2601.02933v3 Announce Type: replace Abstract: Human evaluation is the gold standard for multilingual NLP, but is often skipped in practice and substituted with automatic metrics because it is no

← Previous
1…154155156157158…166
Next →