AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,608 results
27 May 2026

Probabilistic Recurrent Intention Switching Model

ResearchDGX agent

arXiv:2605.26998v1 Announce Type: new Abstract: Inverse reinforcement learning (IRL) recovers reward functions from observed behavior, yet traditional methods assume a single stationary reward that ca

Quoting Kyle Ferrana

ToolsDGX agent

PICARD: Data, shields up DATA: Brilliant! Shields can reduce damage we sustain. Not immunity. Not hubris. Just prudence. It's not precaution—it's strategy. [camera shakes] WORF: HULL BREACHES ON NINE

🚀🚀 Qwen3.7-Max just hit #4 on Code Arena, on par with Claude Opus 4.6 ,top-ranked Chinese lab on the board! @arena More to ship. Stay tune…

Model ReleasesDGX agent

🚀🚀 Qwen3.7-Max just hit #4 on Code Arena, on par with Claude Opus 4.6 ,top-ranked Chinese lab on the board! @arena More to ship. Stay tuned. 🕶️ Qwen3.7 Max (20250517) debuts at #4 in Code Arena: Front

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Stronger models do not always need lighter harnesses. Everyone believes more structured harnesses universally improve reliability, and that …

Model ReleasesDGX agent

Stronger models do not always need lighter harnesses. Everyone believes more structured harnesses universally improve reliability, and that higher-capability models need proportionally less structural

We've created the world's fastest PDF parser ⚡️ And it's more accurate than any other open-source, model-free PDF parser out there (pymupdf,…

Model ReleasesDGX agent

We've created the world's fastest PDF parser ⚡️ And it's more accurate than any other open-source, model-free PDF parser out there (pymupdf, pypdf, markitdown, pdftotext, opendataloader, pymupdf4llm)

Xe-Forge: Multi-Stage LLM-Powered Kernel Optimization for Intel GPU

HardwareDGX agent

arXiv:2605.26118v1 Announce Type: cross Abstract: Porting deep learning algorithms to new hardware accelerators requires developers to repeatedly apply the same low-level optimizations -- quantization

Yohei has the longest running session of any of our users and has been instrumental in our product development process. Super excited to see…

ApplicationsDGX agent

Yohei has the longest running session of any of our users and has been instrumental in our product development process. Super excited to see how activegraph grows! 'If I had Cofounder three years ago,

26 May 2026

After seeing these tweets, I decided to try it out on my own old Ubuntu computer with RTX 1070 GPU (the one that I just upgraded from 16.04 …

HardwareDGX agent

After seeing these tweets, I decided to try it out on my own old Ubuntu computer with RTX 1070 GPU (the one that I just upgraded from 16.04 all the way to 24.04 the other day). Asked Codex on my Mac t

AI-Assisted Systematization for Evaluating GenAI Systems

SafetyDGX agent

arXiv:2605.26001v1 Announce Type: cross Abstract: Evaluating generative AI (GenAI) systems is challenging because many targets of evaluation are broad, contested concepts, such as 'reasoning,' 'fairne

An Interactive Paradigm for Deep Research

Model ReleasesDGX agent

arXiv:2605.24266v1 Announce Type: cross Abstract: Recent advances in large language models (LLMs) have enabled deep research systems that synthesize comprehensive, report-style answers to open-ended q

ATWL: A Formal Language for Representing, Comparing, and Reusing Visual Analytics Workflows

ResearchDGX agent

arXiv:2605.25489v1 Announce Type: new Abstract: Visual analytics (VA) workflows are inherently complex, involving data transformation, feature engineering, visual representation, and human interpretat

Benchmarking and Learning Real-World Customer Service Dialogue

Model ReleasesDGX agent

arXiv:2510.22143v3 Announce Type: replace Abstract: Existing benchmarks and training pipelines for industrial intelligent customer service (ICS) remain misaligned with real-world dialogue requirements

Causal methods for LLM development and evaluation

SafetyDGX agent

arXiv:2605.25998v1 Announce Type: new Abstract: Large language model (LLM) development is currently driven by large-scale empirical iteration over data mixtures, reward models, routing strategies, and

Distilling Game Code World Model Generation into Lightweight Large Language Models

Model ReleasesDGX agent

arXiv:2605.24375v1 Announce Type: new Abstract: Large Language Models (LLMs) have shown great ability in generating executable code from natural language, opening the possibility of automatically cons

Emotional intelligence in large language models is fragmented across perception, cognition, and interaction

Model ReleasesDGX agent

arXiv:2605.24686v1 Announce Type: new Abstract: As large language models (LLMs) are increasingly integrated into emotionally sensitive domains, the structural integrity of their emotional intelligence

FOUND-IT: Foundation-model-first Task-driven 3D Scene Graphs with Granularity on Demand

Model ReleasesDGX agent

arXiv:2605.25371v1 Announce Type: new Abstract: We present the first approach to build hierarchical task-driven 3D scene graphs of arbitrary indoor or outdoor environments using an uncalibrated monocu

From idea to AI app: Creating intelligent research assistants with Strands

IndustryDGX agent

Building an AI app shouldn’t require a PhD in machine learning (ML) or months of wrestling with complex architectures. Yet that’s exactly what happens when you try to orchestrate multiple API calls, m

Grow-Prune-Freeze Networks: Adaptive & Continual Learning Technique for Olfactory Navigation

SafetyDGX agent

arXiv:2605.25170v1 Announce Type: cross Abstract: Training data for olfaction is scattered through disparate, non-standardized datasets that limit the ability to build representative world models. Olf

Huge one for developers building with AI. Excited to announce @ClementDelangue (CEO @huggingface) is joining DASH 2026 for a fireside chat w…

ApplicationsDGX agent

Huge one for developers building with AI. Excited to announce @ClementDelangue (CEO @huggingface) is joining DASH 2026 for a fireside chat with @oliveur. Open source changed software, open models are

Hylos: Operability Contracts for Model-Native Spatial Intelligence

Model ReleasesDGX agent

arXiv:2605.24728v1 Announce Type: new Abstract: Foundation models can increasingly describe, reconstruct, and generate 3D objects, assemblies, scenes, and environments, but visually plausible spatial

Initial benchmarks show Nvidia's Vera CPU, which features 88 in-house-designed Olympus cores, packs a heavy-hitting punch, beating Intel's and AMD's x86_64 CPUs (Michael Larabel/Phoronix)

HardwareDGX agent

Michael Larabel / Phoronix: Initial benchmarks show Nvidia's Vera CPU, which features 88 in-house-designed Olympus cores, packs a heavy-hitting punch, beating Intel's and AMD's x86_64 CPUs — NVIDIA's

LC-ERD: Mining Latent Logic for Self-Evolving Reasoning via Consistency-Regulated Reward Decomposition

SafetyDGX agent

arXiv:2605.24005v1 Announce Type: new Abstract: The evolution of Large Language Model (LLM) reasoning is bottlenecked by the scarcity of high-quality process data. While self-alignment via endogenous

Logic-Guided Socially-aware Robot Navigation World Model

ResearchDGX agent

arXiv:2510.23509v2 Announce Type: replace Abstract: Social robot navigation increasingly relies on large language models for reasoning, path planning, and enabling movement in dynamic human spaces. Ho

Looking forward to speaking at BloombergTech next week with @shiringhaffary to discuss AI safety and our research progress at @LawZero_!

SafetyDGX agent

Looking forward to speaking at BloombergTech next week with @shiringhaffary to discuss AI safety and our research progress at @LawZero_! Is AI development progressing too quickly? @business' @shiringh

Machine Intelligence that Understands Visual and Linguistic Information and Interacts with Humans and Environments

ResearchDGX agent

arXiv:2605.24020v1 Announce Type: cross Abstract: Advancements at the intersection of computer vision and natural language processing are crucial for applications like assistive tech, multimedia query

PEDESTRIANQA: A Benchmark for Vision-Language Models on Pedestrian Intention and Trajectory Prediction

Model ReleasesDGX agent

arXiv:2605.24562v1 Announce Type: cross Abstract: Pedestrian intention and trajectory prediction are critical for the safe deployment of autonomous driving systems, directly influencing navigation dec

Poisoning the Watchtower: Prompt Injection Attacks Against LLM-Augmented Security Operations Through Adversarial Log Content

ResearchDGX agent

arXiv:2605.24421v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used as analyst assistants in security operations centers (SOCs), where they ingest log and alert data t

Right-Sizing Communication and Recommendation Set Size in AI-Assisted Search

SafetyDGX agent

arXiv:2605.23944v1 Announce Type: new Abstract: We model the interaction between a user and an AI driven recommendation system. The user initiates the process by conveying preference information throu

SafeCtrl-RL: Inference-Time Adaptive Behaviour Control for LLM Dialogue via RL-Driven Prompt Optimisation

Model ReleasesDGX agent

arXiv:2605.25984v1 Announce Type: cross Abstract: Ensuring safe and contextually appropriate behaviour in Large Language Models (LLMs) remains a critical challenge for real-world deployment. We presen

Security in the Fine-Tuning Lifecycle of Large Language Models: Threats, Defenses,Evaluation, and Future Directions

Model ReleasesDGX agent

arXiv:2605.25073v1 Announce Type: cross Abstract: Background: Fine-tuning is central to adapting pre-trained Large Language Models (LLMs) to downstream tasks, but its reliance on training data, parame

Teaching large language models to reason like expert diagnosticians

Model ReleasesDGX agent

arXiv:2509.12194v2 Announce Type: replace Abstract: Differential diagnosis is an iterative process that integrates patient information with broader medical knowledge. Clinical case series such as the

The Time is Here for Just-in-Time Systems: Challenges and Opportunities

Model ReleasesDGX agent

arXiv:2605.24096v1 Announce Type: cross Abstract: Core systems like key-value stores have historically taken years to build, and are designed to be general so as to amortize cost across deployments, p

Trust but Verify: Prover-Verifier Deliberation for Selective LLM Prediction

Model ReleasesDGX agent

arXiv:2605.25133v1 Announce Type: new Abstract: Reliably knowing when a language model is correct is almost as important as being correct. We introduce prover-verifier deliberation (PVD), an inference

Truthful Online Preference Aggregation for LLM Fine-Tuning in Mobile Crowdsourcing

Model ReleasesDGX agent

arXiv:2605.24052v1 Announce Type: cross Abstract: To better serve users' demands in mobile applications (e.g., navigation), mobile crowdsourcing platforms can iteratively align large language model (L

Two-Sided Time-Independent Regret for Matching Markets with Limited Interviews

TutorialsDGX agent

arXiv:2602.12224v2 Announce Type: replace-cross Abstract: Two-sided matching platforms rely on preferences from both sides, yet participants can evaluate only a small fraction of potential partners. I

Uncovering Vulnerabilities of LLM-Assisted Cyber Threat Intelligence

ResearchDGX agent

arXiv:2509.23573v4 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly used to help security analysts manage the surge of cyber threats, automating tasks from vulnerab

Unifying Value Alignment and Assignment in Cross-Domain Offline Reinforcement Learning with Heterogeneous Datasets

SafetyDGX agent

arXiv:2605.24862v1 Announce Type: new Abstract: Cross-domain offline reinforcement learning (RL) aims to learn a policy in the target domain with a limited target domain dataset and a source domain da

We aren’t going to do this again so quickly, are we? Rising demand results in higher costs. Higher costs result in lower demand. It is almos…

HardwareDGX agent

We aren’t going to do this again so quickly, are we? Rising demand results in higher costs. Higher costs result in lower demand. It is almost like some sort of equilibrium is being achieved. But there

What Makes a Medical Checker Trainable? Diagnosing Signal Collapse and Reward Hacking in Checker-Guided RAG for Biomedical QA

Model ReleasesDGX agent

arXiv:2605.25988v1 Announce Type: new Abstract: Medical RAG needs evidence-grounded claims, so plugging a claim-level NLI checker into retrieval-augmented RL is intuitive. extbf{We find that the check

25 May 2026

Are Frontier LLMs Ready for Cybersecurity? Evidence for Vertical Foundation Models from Dual-Mode Vulnerability Benchmarks

Model ReleasesDGX agent

arXiv:2605.23243v1 Announce Type: cross Abstract: We evaluate whether frontier LLMs are ready for cybersecurity through a dual-mode benchmark: white-box function-level vulnerability detection (VulnLLM

DepthAgent: Towards Better Universal Depth Estimation via Sample-wise Expert Selection

Model ReleasesDGX agent

arXiv:2605.23281v1 Announce Type: new Abstract: Monocular metric depth estimation has achieved strong progress with large-scale training and universal-camera modeling, yet robust deployment across div

Direct Dynamic Retargeting for Humanoid Imitation Learning from Videos

SafetyDGX agent

arXiv:2605.23762v1 Announce Type: new Abstract: Imitation Learning from monocular video demonstrations provides a scalable approach for teaching complex skills to humanoid robots. However, translating

DRL-Driven Edge-Aware Utility Optimization for Multi-Slice 6G Networks

ResearchDGX agent

arXiv:2605.23056v1 Announce Type: cross Abstract: Virtual Reality (VR) services delivered over 6G networks demand ultra-low latency and high bandwidth to ensure seamless user experiences. This paper p

Evaluating Large Language Models in a Complex Hidden Role Game

Model ReleasesDGX agent

arXiv:2605.22826v1 Announce Type: cross Abstract: Quantifying the deceptive potential of Large Language Models (LLMs) is critical for AI safety, yet difficult to achieve in uncontrolled environments.

Fast-dDrive: Efficient Block-Diffusion VLM for Autonomous Driving

SafetyDGX agent

arXiv:2605.23163v1 Announce Type: new Abstract: End-to-end autonomous driving via Vision-Language-Action (VLA) models demands a precarious balance between high-fidelity trajectory planning and efficie

How Far Are We from Generating Missing Modalities with Foundation Models?

Model ReleasesDGX agent

arXiv:2506.03530v3 Announce Type: replace-cross Abstract: Multimodal foundation models have demonstrated impressive capabilities across diverse tasks. However, their potential as plug-and-play solutio

Positional Failures in Long-Context LLMs: A Blind Spot in Reasoning Benchmarks

Model ReleasesDGX agent

arXiv:2605.23170v1 Announce Type: cross Abstract: Position-controlled evaluation is standard for retrieval tasks such as Needle-in-a-Haystack and RULER, but mainstream reasoning benchmarks do not cont

Searching X in Grok Build is really fast and useful for research/docs/real-time info, video is not sped up

IndustryDGX agent

Searching X in Grok Build is really fast and useful for research/docs/real-time info, video is not sped up Media I've been using Grok Build some time now, and it has genuinely surprised me. It is uniq

SeedER: Seed-and-Expand Retrieval from Knowledge Graphs

SafetyDGX agent

arXiv:2605.23753v1 Announce Type: new Abstract: Knowledge graphs (KGs) offer a rich representation for relational knowledge, but their irregular structure makes retrieval challenging: ego-graph expans

Strategic Coercion Within Alliances: The Greenland Sovereignty Game as an AI Stress Test

Model ReleasesDGX agent

arXiv:2605.22841v1 Announce Type: cross Abstract: What happens when the strongest alliance member pressures a weaker member over territory and strategic control? We examine the Greenland sovereignty c

The Deterministic Horizon: Impossibility Results as Design Specifications for Trustworthy AI Systems

ApplicationsDGX agent

arXiv:2605.23024v1 Announce Type: new Abstract: Large language models now write software, draft legal documents, and produce clinical notes, yet fundamental limits, from Turing and Arrow to the No Fre

// The Efficiency Frontier in LLMs // (bookmark this one) How much are you overpaying for context you do not need? It turns out that context…

TutorialsDGX agent

// The Efficiency Frontier in LLMs // (bookmark this one) How much are you overpaying for context you do not need? It turns out that context costs dominate production LLM bills, and the right strategy

USIM and U0: A Vision-Language-Action Dataset and Model for General Underwater Robots

ResearchDGX agent

arXiv:2510.07869v4 Announce Type: replace Abstract: Underwater environments pose unique challenges for robotic navigation and manipulation. While existing research has primarily focused on task-specif

24 May 2026

⚠️⚠️⚠️i don’t think most people understand the implications of the mood shift below, so I will spell them out. they are serious, and eventua…

SafetyDGX agent

⚠️⚠️⚠️i don’t think most people understand the implications of the mood shift below, so I will spell them out. they are serious, and eventually will affect the global economy. when a serious coder as

neurosymbolic by @swarat et al for the Erdos win, with much more careful, quantitative work than openai’s in hindsight i wonder whether Open…

SafetyDGX agent

neurosymbolic by @swarat et al for the Erdos win, with much more careful, quantitative work than openai’s in hindsight i wonder whether OpenAI rushed theirs out, knowing this was coming? Another 9 ope

Quoting Armin Ronacher

ToolsDGX agent

The most frustrating failure mode right now is that people submit issues that are not in their own voice. They contain an observed problem somewhere, but it has been thrown into a clanker and the clan

“We’re focused on providing the best in class document understanding and OCR technologies.” In partnership with @Wing_VC's 2026 Enterprise T…

ApplicationsDGX agent

“We’re focused on providing the best in class document understanding and OCR technologies.” In partnership with @Wing_VC's 2026 Enterprise Tech 30, @jerryjliu0, Co-Founder & CEO of @Llama_Index, visit

23 May 2026

CCLab: Adversarial Testing of Learning- and Non-Learning-Based Congestion Controllers

SafetyDGX agent

arXiv:2605.21915v1 Announce Type: cross Abstract: Congestion controllers (CCs) are critical to network performance, and yet their robustness under adverse conditions remains insufficiently understood.

Chebyshev Policies and the Mountain Car Problem: Reinforcement Learning for Low-Dimensional Control Tasks

Model ReleasesDGX agent

arXiv:2605.22305v1 Announce Type: new Abstract: We analytically solve the Mountain Car problem, a canonical benchmark in RL, and derive an optimal control solution, closing a gap after 36 years. This

Long-term Fairness with Selective Labels

SafetyDGX agent

arXiv:2605.22291v1 Announce Type: new Abstract: Long-term fairness algorithms aim to satisfy fairness beyond static and short-term notions by accounting for the dynamics between decision-making polici

← Previous
1…275276277278279…294
Next →