AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,599 results
14 Apr 2026

One of my passions is that education should be dispersed freely and as widely as possible, especially for technologies as dynamic and crucia…

Model ReleasesDGX agent

One of my passions is that education should be dispersed freely and as widely as possible, especially for technologies as dynamic and crucial as LLMs/AI. I'm proud to have friends who would disown me

Online Learning-Enhanced High Order Adaptive Safety Control

SafetyDGX agent

arXiv:2511.19651v2 Announce Type: replace Abstract: Control barrier functions (CBFs) are an effective model-based tool to formally certify the safety of a system. With the growing complexity of modern

OOWM: Structuring Embodied Reasoning and Planning via Object-Oriented Programmatic World Modeling

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.09580v1 Announce Type: new Abstract: Standard Chain-of-Thought (CoT) prompting empowers Large Language Models (LLMs) with reasoning capabilities, yet its reliance on linear natural language

OpenAI partners with Novo Nordisk to accelerate drug discovery and delivery

IndustryDGX agent

Artificial intelligence powerhouse OpenAI Group PBC is lending its expertise to Novo Nordisk A/S in a new partnership announced today that aims to accelerate research in the pharmaceutical industry. T

Openclaw with Gemma4 26B extremely slow and forget stuff

Local AiDGX agent

Users in the r/ollama community report that running OpenClaw with the Gemma4 26B model via Ollama results in pathologically slow first-turn performance, with the degree of slowdown scaling with OpenCl

Pando: Do Interpretability Methods Work When Models Won't Explain Themselves?

Model ReleasesDGX agent

arXiv:2604.11061v1 Announce Type: cross Abstract: Mechanistic interpretability is often motivated for alignment auditing, where a model's verbal explanations can be absent, incomplete, or misleading.

Policy-Guided Threat Hunting: An LLM enabled Framework with Splunk SOC Triage

Model ReleasesDGX agent

arXiv:2603.23966v3 Announce Type: replace-cross Abstract: With frequently evolving Advanced Persistent Threats (APTs) in cyberspace, traditional security solutions approaches have become inadequate fo

Preference-Agile Multi-Objective Optimization for Real-time Vehicle Dispatching

SafetyDGX agent

arXiv:2604.10664v1 Announce Type: new Abstract: Multi-objective optimization (MOO) has been widely studied in literature because of its versatility in human-centered decision making in real-life appli

Pyramid MoA: A Probabilistic Framework for Cost-Optimized Anytime Inference

ResearchDGX agent

arXiv:2602.19509v3 Announce Type: replace-cross Abstract: We observe that LLM cascading and routing implicitly solves an anytime computation problem -- a class of algorithms, well-studied in classical

Radiology Report Generation for Low-Quality X-Ray Images

Model ReleasesDGX agent

arXiv:2604.10188v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have significantly advanced automated Radiology Report Generation (RRG). However, existing methods implicitly assume high-

Rethinking LLM Watermark Detection in Black-Box Settings: A Non-Intrusive Third-Party Framework

SafetyDGX agent

arXiv:2603.14968v2 Announce Type: replace-cross Abstract: While watermarking serves as a critical mechanism for LLM provenance, existing secret-key schemes tightly couple detection with injection, req

RiTeK: A Dataset for Large Language Models Complex Reasoning over Textual Knowledge Graphs in Medicine

Model ReleasesDGX agent

arXiv:2410.13987v3 Announce Type: replace Abstract: Answering complex real-world questions in the medical domain often requires accurate retrieval from medical Textual Knowledge Graphs (medical TKGs),

Robust Real-Time Coordination of CAVs: A Distributed Optimization Framework under Uncertainty

SafetyDGX agent

arXiv:2508.21322v2 Announce Type: replace Abstract: Achieving both safety guarantees and real-time performance in cooperative vehicle coordination remains a fundamental challenge, particularly in dyna

See Fair, Speak Truth: Equitable Attention Improves Grounding and Reduces Hallucination in Vision-Language Alignment

SafetyDGX agent

arXiv:2604.09749v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) frequently hallucinate objects that are absent from the visual input, often because attention during decoding i

Self-Evolving LLM Memory Extraction Across Heterogeneous Tasks

Model ReleasesDGX agent

arXiv:2604.11610v1 Announce Type: new Abstract: As LLM-based assistants become persistent and personalized, they must extract and retain useful information from past conversations as memory. However,

SHE: Stepwise Hybrid Examination Reinforcement Learning Framework for E-commerce Search Relevance

SafetyDGX agent

arXiv:2510.07972v3 Announce Type: replace Abstract: Query-product relevance prediction is vital for AI-driven e-commerce, yet current LLM-based approaches face a dilemma: SFT and DPO struggle with lon

Sources: Microsoft ended production of its Surface Hub 3 collaborative touch displays and scrapped plans for a Hub 4; Hub 3 debuted in 2023 in 50' and 85' sizes (Zac Bowden/Windows Central)

ApplicationsDGX agent

Zac Bowden / Windows Central: Sources: Microsoft ended production of its Surface Hub 3 collaborative touch displays and scrapped plans for a Hub 4; Hub 3 debuted in 2023 in 50' and 85' sizes — Microso

The @MiniMax_AI team is receptive to community feedback and are updating the license. See the latest here. https://x.com/RyanLeeMiniMax/stat…

Local AiDGX agent

The @MiniMax_AI team is receptive to community feedback and are updating the license. See the latest here. https://x.com/RyanLeeMiniMax/status/2044132777877221515?s=20 I just updated our license. For

Thought Branches: Interpreting LLM Reasoning Requires Resampling

SafetyDGX agent

arXiv:2510.27484v2 Announce Type: replace-cross Abstract: Most work interpreting reasoning models studies only a single chain-of-thought (CoT), yet these models define distributions over many possible

TimeSeriesExamAgent: Creating Time Series Reasoning Benchmarks at Scale

Model ReleasesDGX agent

arXiv:2604.10291v1 Announce Type: new Abstract: Large Language Models (LLMs) have shown promising performance in time series modeling tasks, but do they truly understand time series data? While multip

Towards Adaptive Open-Set Object Detection via Category-Level Collaboration Knowledge Mining

SafetyDGX agent

arXiv:2604.11195v1 Announce Type: cross Abstract: Existing object detectors often struggle to generalize across domains while adapting to emerging novel categories. Adaptive open-set object detection

Tuning Qwen2.5-VL to Improve Its Web Interaction Skills

Model ReleasesDGX agent

arXiv:2604.09571v1 Announce Type: cross Abstract: Recent advances in vision-language models (VLMs) have sparked growing interest in using them to automate web tasks, yet their feasibility as independe

Valence-Arousal Subspace in LLMs: Circular Emotion Geometry and Multi-Behavioral Control

Model ReleasesDGX agent

arXiv:2604.03147v2 Announce Type: replace-cross Abstract: We present a method to identify a valence-arousal (VA) subspace within large language model representations. From 211k emotion-labeled texts,

We achieved state-of-the-art performance in predicting which of 4.2 million genetic variants cause diseases by interpreting a genomics model…

ToolsDGX agent

We achieved state-of-the-art performance in predicting which of 4.2 million genetic variants cause diseases by interpreting a genomics model, in a new preprint with @MayoClinic. We're now releasing an

WebLLM: A High-Performance In-Browser LLM Inference Engine

Local AiDGX agent

arXiv:2412.15803v2 Announce Type: replace-cross Abstract: Advancements in large language models (LLMs) have unlocked remarkable capabilities. While deploying these models typically requires server-gra

What I’ve been building: ATOM Report, post-training course, finishing my book, and ongoing research

TutorialsDGX agent

The author provided an update on several ongoing technical projects, including the ATOM Report, a new post-training course, and

What to Say and When to Say it: Live Fitness Coaching as a Testbed for Situated Interaction

Model ReleasesDGX agent

arXiv:2407.08101v4 Announce Type: replace Abstract: Vision-language models have shown impressive progress in recent years. However, existing models are largely limited to turn-based interactions, wher

When Can You Poison Rewards? A Tight Characterization of Reward Poisoning in Linear MDPs

SafetyDGX agent

arXiv:2604.10062v1 Announce Type: new Abstract: We study reward poisoning attacks in reinforcement learning (RL), where an adversary manipulates rewards within constrained budgets to force the target

Who Gets Which Message? Auditing Demographic Bias in LLM-Generated Targeted Text

Model ReleasesDGX agent

arXiv:2601.17172v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly capable of generating personalized, persuasive text at scale, raising new questions about bias a

Why Smaller Is Slower? Dimensional Misalignment in Compressed LLMs

Model ReleasesDGX agent

arXiv:2604.09595v1 Announce Type: cross Abstract: Post-training compression reduces LLM parameter counts but often produces irregular tensor dimensions that degrade GPU performance -- a phenomenon we

ZoomR: Memory Efficient Reasoning through Multi-Granularity Key Value Retrieval

SafetyDGX agent

arXiv:2604.10898v1 Announce Type: new Abstract: Large language models (LLMs) have shown great performance on complex reasoning tasks but often require generating long intermediate thoughts before reac

13 Apr 2026

$200/month is enough to buy an H100 GPU for 6 hours every workday

HardwareDGX agent

Soumith Chintala shared a post highlighting that $200 per month is sufficient to rent access to an NVIDIA H100 GPU for approximately 6 hours every workday, making high-end AI compute more accessible t

Advantage-Guided Diffusion for Model-Based Reinforcement Learning

SafetyDGX agent

arXiv:2604.09035v1 Announce Type: new Abstract: Model-based reinforcement learning (MBRL) with autoregressive world models suffers from compounding errors, whereas diffusion world models mitigate this

ALMAB-DC: Active Learning, Multi-Armed Bandits, and Distributed Computing for Sequential Experimental Design and Black-Box Optimization

Model ReleasesDGX agent

arXiv:2603.21180v3 Announce Type: replace Abstract: Sequential experimental design under expensive, gradient-free objectives is a central challenge in computational statistics: evaluation budgets are

AniGen: Unified S^3 Fields for Animatable 3D Asset Generation

ApplicationsDGX agent

arXiv:2604.08746v1 Announce Type: cross Abstract: Animatable 3D assets, defined as geometry equipped with an articulated skeleton and skinning weights, are fundamental to interactive graphics, embodie

Another great AiE in the books, this time in 🇪🇺 Europe for Arize AI and @arizephoenix So great to see all the homies again and learn a few…

TutorialsDGX agent

Another great AiE in the books, this time in 🇪🇺 Europe for Arize AI and @arizephoenix So great to see all the homies again and learn a few things myself in one of my favorite cities Great themes this

AVA-VLA: Improving Vision-Language-Action models with Active Visual Attention

SafetyDGX agent

arXiv:2511.18960v3 Announce Type: replace-cross Abstract: Vision-Language-Action (VLA) models have shown remarkable progress in embodied tasks recently, but most methods process visual observations in

Balancing User Preferences by Social Networks: A Condition-Guided Social Recommendation Model for Mitigating Popularity Bias

SafetyDGX agent

arXiv:2405.16772v2 Announce Type: replace-cross Abstract: Social recommendation models weave social interactions into their design to provide uniquely personalized recommendation results for users. Ho

CausalVAD: De-confounding End-to-End Autonomous Driving via Causal Intervention

SafetyDGX agent

arXiv:2603.18561v2 Announce Type: replace Abstract: Planning-oriented end-to-end driving models show great promise, yet they fundamentally learn statistical correlations instead of true causal relatio

Chain-in-Tree: Back to Sequential Reasoning in LLM Tree Search

SafetyDGX agent

arXiv:2509.25835v4 Announce Type: replace Abstract: Test-time scaling improves large language models (LLMs) on long-horizon reasoning tasks by allocating more compute at inference. LLM inference via t

Codex with Voiden

Local AiDGX agent

'Voiden' doesn't appear in any search results as a known model or tool in the Ollama ecosystem. Based on the Reddit source and the broader context of the r/ollama community, this post likely discusses

Detection and Characterization of Coordinated Online Behavior: A Survey

TutorialsDGX agent

arXiv:2408.01257v2 Announce Type: replace-cross Abstract: Coordination is a fundamental aspect of life. The advent of social media has made it integral also to online human interactions, such as those

E3-TIR: Enhanced Experience Exploitation for Tool-Integrated Reasoning

SafetyDGX agent

arXiv:2604.09455v1 Announce Type: new Abstract: While Large Language Models (LLMs) have demonstrated significant potential in Tool-Integrated Reasoning (TIR), existing training paradigms face signific

EgoTL: Egocentric Think-Aloud Chains for Long-Horizon Tasks

Model ReleasesDGX agent

arXiv:2604.09535v1 Announce Type: new Abstract: Large foundation models have made significant advances in embodied intelligence, enabling synthesis and reasoning over egocentric input for household ta

EmoCtrl: Controllable Emotional Image Content Generation

SafetyDGX agent

arXiv:2512.22437v2 Announce Type: replace Abstract: An image conveys meaning through both its visual content and emotional tone, jointly shaping human perception. We introduce Controllable Emotional I

Fine-tuning is Not Enough: A Parallel Framework for Collaborative Imitation and Reinforcement Learning in End-to-end Autonomous Driving

Model ReleasesDGX agent

arXiv:2603.13842v3 Announce Type: replace-cross Abstract: End-to-end autonomous driving is typically built upon imitation learning (IL), yet its performance is constrained by the quality of human demo

FIT-GNN: Faster Inference Time for GNNs that 'FIT' in Memory Using Coarsening

Model ReleasesDGX agent

arXiv:2410.15001v5 Announce Type: replace Abstract: Scalability of Graph Neural Networks (GNNs) remains a significant challenge. To tackle this, methods like coarsening, condensation, and computation

from my experience, even the best models (Opus 4.6, 5.4 xhigh / 5.3 codex) cannot write good code today without an amount of work that is eq…

TutorialsDGX agent

from my experience, even the best models (Opus 4.6, 5.4 xhigh / 5.3 codex) cannot write good code today without an amount of work that is equivalent to just doing the work myself am excited for a worl

Is an nvidia DGK Spark or similar worth it?

HardwareDGX agent

This Reddit thread on r/ollama discusses whether the NVIDIA DGX Spark — powered by the GB10 Grace Blackwell Superchip and delivering 1 petaFLOP of performance — is a worthwhile investment for running

LEGO: Latent-space Exploration for Geometry-aware Optimization of Humanoid Kinematic Design

TutorialsDGX agent

arXiv:2604.08636v1 Announce Type: cross Abstract: Designing robot morphologies and kinematics has traditionally relied on human intuition, with little systematic foundation. Motion-design co-optimizat

Listener-Rewarded Thinking in VLMs for Image Preferences

Model ReleasesDGX agent

arXiv:2506.22832v3 Announce Type: replace-cross Abstract: Training robust and generalizable reward models for human visual preferences is essential for aligning text-to-image and text-to-video generat

LLM Dictionary: A reference to contemporary LLM vocabulary [P]

ResearchDGX agent

This Reddit post on r/MachineLearning presents a community-contributed dictionary of contemporary Large Language Model (LLM) terminology, covering terms related to training, fine-tuning, inference, al

LMGenDrive: Bridging Multimodal Understanding and Generative World Modeling for End-to-End Driving

SafetyDGX agent

arXiv:2604.08719v1 Announce Type: cross Abstract: Recent years have seen remarkable progress in autonomous driving, yet generalization to long-tail and open-world scenarios remains a major bottleneck

Mamba-Based Graph Convolutional Networks: Tackling Over-smoothing with Selective State Space

Model ReleasesDGX agent

arXiv:2501.15461v4 Announce Type: replace Abstract: Graph Neural Networks (GNNs) have shown great success in various graph-based learning tasks. However, it often faces the issue of over-smoothing as

MedConceal: A Benchmark for Clinical Hidden-Concern Reasoning Under Partial Observability

Model ReleasesDGX agent

arXiv:2604.08788v1 Announce Type: new Abstract: Patient-clinician communication is an asymmetric-information problem: patients often do not disclose fears, misconceptions, or practical barriers unless

Nvidia hasn’t budged in six months. Coreweave is down 21% Oracle is down 50% Microsoft is down 25% Like it or not, the bubble has already st…

HardwareDGX agent

Gary Marcus, an AI skeptic and cognitive scientist, posted on X highlighting diverging stock performance in the AI infrastructure sector, noting that while Nvidia has remained relatively stable, CoreW

ParseBench is here!📊 We’ve just released ParseBench, an open benchmark + dataset for evaluating document parsing at scale. It includes: • 2…

Model ReleasesDGX agent

ParseBench is here!📊 We’ve just released ParseBench, an open benchmark + dataset for evaluating document parsing at scale. It includes: • 2,000+ human-reviewed enterprise documents • 167,000 evaluatio

Policy-Aware Design of Large-Scale Factorial Experiments

SafetyDGX agent

arXiv:2604.08804v1 Announce Type: cross Abstract: Digital firms routinely run many online experiments on shared user populations. When product decisions are compositional, such as combinations of inte

RAMP: Hybrid DRL for Online Learning of Numeric Action Models

SafetyDGX agent

arXiv:2604.08685v1 Announce Type: new Abstract: Automated planning algorithms require an action model specifying the preconditions and effects of each action, but obtaining such a model is often hard.

Robust Adaptive Backstepping Impedance Control of Robots in Unknown Environments

ResearchDGX agent

arXiv:2604.09323v1 Announce Type: new Abstract: This paper presents a Robust Adaptive Backstepping Impedance Control (RABIC) strategy for robots operating in contact-rich and uncertain environments. T

← Previous
1…289290291292293294
Next →