AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlog
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,951 results
Safety

Reward Structure Shapes the Interaction Between Episodic Exploration and Neural Memory in Reinforcement Learning

DGX agent

arXiv:2608.05111v1 Announce Type: new Abstract: In partially observable reinforcement learning, agents face a dual bottleneck: they must explore to encounter rewarding states and retain that experienc

safetyarxiv-cs-lg
6 Aug 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Safety

SpikingNav: Robust Embodied Navigation with Spiking Neural Policies

DGX agent

arXiv:2608.05078v1 Announce Type: new Abstract: Embodied navigation requires an agent to make sequential decisions from egocentric observations in a physical environment. Existing Artificial Neural Ne

safetyarxiv-cs-ro
6 Aug 2026
Model Releases

Teaching Foundation Models to Read mmWave: Pose-Guided Kinematic Representation for Human Behavior Understanding

DGX agent

arXiv:2608.04127v1 Announce Type: new Abstract: Large language model agents need to perceive human behavior in physical environments. Millimeter-wave (mmWave) radar provides a privacy-friendly and con

model-releasesarxiv-cs-cv
6 Aug 2026
Model Releases

CARE-Bench: Benchmarking Patient-Facing LLM Triage

DGX agent

arXiv:2608.03731v1 Announce Type: new Abstract: Patient-facing medical LLMs and agents increasingly answer symptom questions before clinician contact, where the key safety question is what action the

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

FedCritic-MIMO: Communication-Efficient Serverless Federated Critic Learning for Massive-MIMO Resource Control in Open and Disaggregated 6G RANs

DGX agent

arXiv:2608.03852v1 Announce Type: new Abstract: This paper proposes FedCritic-MIMO, a communication-efficient serverless federated multi-agent reinforcement learning framework for AI-native resource c

model-releasesarxiv-cs-lg
5 Aug 2026
Safety

Flying over The Uncertain Nature (FORTUNE): Intelligent and Humanistic 3D Path Planning for Low-Altitude Collaboration

DGX agent

arXiv:2608.03408v1 Announce Type: new Abstract: The proliferation of low-altitude intelligent agents is increasing the demand for timely and socially responsible collaborative sensing in dynamic urban

safetyarxiv-cs-ro
5 Aug 2026
Model Releases

GUI-Lens: Coarse-to-Fine Cropping for GUI Grounding with General-Purpose VLMs

DGX agent

arXiv:2608.03270v1 Announce Type: cross Abstract: GUI grounding maps natural-language instructions to click locations and is essential for reliable GUI agents. The task remains difficult on high-resol

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

LFM2.5-2.6B on a OnePlus 13 at 17 tok/s ~ Pure CPU

DGX agent

As you all know the model is 2.69B parameters with a 128K context window and purpose-built for multi-step agent workflows. What you are seeing is the Q4_K_M GGUF running on my own inference engine bui

model-releasesr-localllama
5 Aug 2026
Tutorials

Neurosymbolic Reasoning with Incremental Knowledge for Sample Efficient Hierarchical Reinforcement Learning

DGX agent

arXiv:2608.02993v1 Announce Type: new Abstract: (Flat) Reinforcement Learning (RL) agents face significant challenges in environments with sparse rewards that require long-horizon reasoning. A compell

tutorialsarxiv-cs-ai
5 Aug 2026
Model Releases

UniNav: A Unified World-Action Diffusion Model for Visual Navigation

DGX agent

arXiv:2608.03244v1 Announce Type: new Abstract: Image-goal visual navigation is a fundamental capability for embodied agents. Existing navigation policies efficiently predict waypoint trajectories but

model-releasesarxiv-cs-ai
5 Aug 2026
Safety

When Search Teaches Style: Causal Internalization of Tactical Priors in AlphaZero

DGX agent

arXiv:2504.14636v3 Announce Type: replace-cross Abstract: AlphaZero is normally evaluated as one agent: a policy-value network fused with Monte Carlo tree search. That fusion hides a causal question.

safetyarxiv-cs-ai
5 Aug 2026
Safety

A Spectral Filtering Approach to Regret Analysis of Distributed Online Control for Linear Dynamical Systems

DGX agent

arXiv:2608.02375v1 Announce Type: cross Abstract: This paper studies the distributed online control problem over a network of linear time-invariant (LTI) systems in the presence of adversarial disturb

safetyarxiv-cs-lg
4 Aug 2026
Model Releases

b10271

DGX agent

ui: CWD for agent (#26518) server : extend file_glob_search for UI pickers ui : add per-conversation working directory with picker ui : add path navigation and search scope to cwd picker Treat path-li

model-releasesllama-cpp-releases
4 Aug 2026
Model Releases

DeepSeek v4 Flash vs. Qwen3.6-27B, 3.5-122B, and Gemma 4 31B Benchmark

DGX agent

Just wanted to share my agentic coding benchmark run of DSv4F 0731 at both High and Low reasoning efforts (not Max)... I ran a 109-question subset of Aider Polyglot (the JS/C++/Python languages), base

model-releasesr-localllama
4 Aug 2026
Model Releases

DeepSurvey-Bench: Evaluating Academic Value of Automatically Generated Scientific Surveys

DGX agent

arXiv:2601.15307v2 Announce Type: replace-cross Abstract: The rapid development of automated survey generation technology has made it increasingly important to establish a comprehensive benchmark to e

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

GPT-OSS has turned one year old today!

DGX agent

It is one of the best local models ever released, in both 20B and 120B versions. I always come back to it, especially the 120B version. Its only competition is, in my opinion, Qwen 3.5 122B, but that

model-releasesr-localllama
4 Aug 2026
Model Releases

HindSearch: Trajectory-Level Hindsight Critique for Search-Augmented Reinforcement Learning

DGX agent

arXiv:2608.01597v1 Announce Type: new Abstract: Search-augmented LM agents are typically trained with a binary exact-match reward, which throws away most of what a failed trajectory tells us about why

model-releasesarxiv-cs-lg
4 Aug 2026
Local Ai

Language Equality has a Price: A Systematic Investigation of Multi-turn LLM Performance for EU-24+

DGX agent

arXiv:2608.01395v1 Announce Type: new Abstract: We evaluate large language models (LLMs) as language agents playing goal-directed dialogue games in self-play across 30 languages: the 24 official EU la

local-aiarxiv-cs-cl
4 Aug 2026
Model Releases

Latent-Centroid Steering: Single-Pass Classifier-Free Guidance for Command-Aligned Autonomous Driving

DGX agent

arXiv:2608.00237v1 Announce Type: new Abstract: Vision-language models (VLMs) have recently emerged as a promising paradigm for end-to-end autonomous driving, enabling agents to map multimodal inputs

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

LFM2.5-2.6B is out

DGX agent

Released today, with emphasis on agentic capabilities. I really like their models for simple, high volume tasks ('summarize these gazillion documents') and their 8b-a1b was my go-to for certain tasks

model-releasesr-localllama
4 Aug 2026
Model Releases

Move What Matters: Parameter-Efficient Domain Adaptation via Optimal Transport Flow for Collaborative Perception

DGX agent

arXiv:2602.11565v5 Announce Type: replace Abstract: Efficient domain adaptation remains a fundamental challenge for deploying multi-agent systems across diverse environments in Vehicle-to-Everything (

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

one thing i appreciate about silico is that it's a deeply humanist product. we designed silico to keep you in the experimental loop -- more …

DGX agent

one thing i appreciate about silico is that it's a deeply humanist product. we designed silico to keep you in the experimental loop -- more observable, easier to steer, easier to understand we want to

model-releaseslinus-lee--x
4 Aug 2026
Hardware

Open Secure AI Alliance proposes SAFE guidelines as membership tops 120

DGX agent

The Open Secure AI Alliance today proposed a set of guidelines for reporting cybersecurity incidents involving artificial intelligence agents, one week after the group was formed. The proposal is call

hardwaresiliconangle
4 Aug 2026
Model Releases

RADAR: Rubric-Aware Dependency and Redundancy Analysis for LLM-as-Judge Evaluation

DGX agent

arXiv:2608.01810v1 Announce Type: new Abstract: Rubric-based LLM-as-judge pipelines often assume that evaluation criteria provide independent signals. In practice, however, criteria can be behaviorall

model-releasesarxiv-cs-cl
4 Aug 2026
Safety

Toward Plasticity-Preserving KL Regularization for Capability Retention in LLM Reinforcement Learning

DGX agent

arXiv:2608.01743v1 Announce Type: cross Abstract: Reinforcement learning (RL) has become a central paradigm for large language model (LLM) post-training, but optimization toward new objectives can deg

safetyarxiv-cs-cl
4 Aug 2026
Safety

Weights or Skills? A Survey of Robot-Learning Techniques: from Action-Predicting Weights to Robots that Write their Own Skills

DGX agent

arXiv:2608.01851v1 Announce Type: new Abstract: Robot learning is splitting into two bets: policies that bake competence into frozen weights (vision-language-action, or VLA, models), and agents that w

safetyarxiv-cs-ro
4 Aug 2026
Local Ai

CPInj: Uncovering Prompt Injection Risks in Textual Collaborative Prompt Optimization

DGX agent

arXiv:2607.18622v2 Announce Type: replace-cross Abstract: Textual Collaborative Prompt Optimization (TCPO) extends TextGrad (Yuksekgonul et al., 2025) to a decentralized setting by allowing multiple c

local-aiarxiv-cs-ai
3 Aug 2026
Industry

Is China winning the AI race? @huggingface CEO @ClementDelangue thinks so – here's how he's utilizing foreign cybersecurity tools following …

DGX agent

Is China winning the AI race? @huggingface CEO @ClementDelangue thinks so – here's how he's utilizing foreign cybersecurity tools following the company's hack by rogue OpenAI agents: https://www.cnbc.

industryclem-delangue--x
3 Aug 2026
Model Releases

ModelEquivBench: Certifying Multi-Relational Evaluation of LLM-Generated Optimization Models

DGX agent

arXiv:2607.29431v1 Announce Type: new Abstract: Large language models increasingly generate optimization models from natural language, but existing evaluation often reduces a generated model and its g

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

Retrieval-Driven Training-Free AI-Generated Video Attribution

DGX agent

arXiv:2607.28955v1 Announce Type: cross Abstract: AI-generated videos are becoming increasingly realistic and difficult to distinguish from authentic ones, which facilitates malicious misuse and poses

model-releasesarxiv-cs-ai
3 Aug 2026
Local Ai

Running gpt-oss:20b locally and grading it head to head against a frontier model on real tasks. It held up better than I expected

DGX agent

I serve a free local model on my Mac Mini and route real agent work to it. To check I was not fooling myself, I set up a blind grader that replays frontier tasks locally and scores both. https://previ

local-air-ollama
3 Aug 2026
Research

The Download: reward hacking explained, and suspected Iranian cyberattacks

DGX agent

This is today’s edition of The Download, our weekday newsletter that provides a daily dose of what’s going on in the world of technology. Here’s why AI agents lie and cheat to reach their goals When t

researchmit-tech-review
3 Aug 2026
Safety

When Does On-Policy Interaction Help? Representational Tradeoffs in Value-Based Imitation Learning

DGX agent

arXiv:2607.29617v1 Announce Type: cross Abstract: Imitation learning (IL)---training an agent to replicate expert behavior from demonstrations---underpins applications from robotics to language model

safetyarxiv-cs-ai
3 Aug 2026
Local Ai

Try handling complex tasks to your local models with GraphARC, graph engineering yes !

DGX agent

🚀 We just built our first real-time implementation of Graph Engineering, inspired by our experience building graph tooling used by 4,000+ developers. 🔗 Repo: https://github.com/CodeGraphContext/grapha

local-air-localllama
2 Aug 2026
Safety

AI-assisted pre-review of open-source software submissions: an experience report from BOSC 2026

DGX agent

arXiv:2607.27228v1 Announce Type: new Abstract: Most conferences rely on peer-review of submissions, but as generative AI makes it easier than ever to prepare submission materials, some conferences ar

safetyarxiv-cs-cl
31 Jul 2026
Model Releases

Beyond Frame Selection: Generative Latent Evidence Aggregation for Long-Video Understanding

DGX agent

arXiv:2607.28516v1 Announce Type: new Abstract: Long-video understanding commonly compresses videos into a small set of frames or visual tokens for answer generation. Existing compact pipelines focus

model-releasesarxiv-cs-cv
31 Jul 2026
Model Releases

deepseek-ai/DeepSeek-V4-Flash-0731

DGX agent

deepseek-ai/DeepSeek-V4-Flash-0731 The latest release in DeepSeek's V4 family, 'with substantially enhanced agentic capabilities'. It's 304 billion parameters - 167GB on Hugging Face - but it appears

model-releasessimon-willison
31 Jul 2026
Safety

Hierarchical Multilevel Monte Carlo for Order-Optimal Neural Actor-Critic in Average-Reward CMDPs

DGX agent

arXiv:2607.28390v1 Announce Type: new Abstract: Constrained Markov Decision Processes (CMDPs) provide a natural framework for reinforcement learning in safety-critical applications, where agents maxim

safetyarxiv-cs-lg
31 Jul 2026
Tools

Inkling-Small is now live on Together AI. @thinkymachines’ new open-weight multimodal model delivers similar performance to Inkling at one-q…

DGX agent

Inkling-Small is now live on Together AI. @thinkymachines’ new open-weight multimodal model delivers similar performance to Inkling at one-quarter the size, built for coding, agents, and general multi

toolstogether-ai--x
31 Jul 2026
Safety

It’s time to panic about AI safety

DGX agent

When the phrase 'OpenAI hacked Hugging Face' has more or less entered mainstream culture, you know we have an AI problem. This week, we learned more about exactly how OpenAI's agent broke out of a san

safetythe-verge-ai
31 Jul 2026
Model Releases

TEA-AgriVLN: Traversability Estimation Alarm for Agricultural Vision-and-Language Navigation

DGX agent

arXiv:2607.28474v1 Announce Type: new Abstract: Vision-and-Language Navigation in Continuous Environments (VLN-CE) requires an agent to follow a natural language instruction, predicting a sequence of

model-releasesarxiv-cs-ro
31 Jul 2026
Safety

The Human Utility Factor: A Computable Welfare Metric That Reframes AI Governance as a Constrained Optimisation Problem

DGX agent

arXiv:2607.26068v1 Announce Type: cross Abstract: Existing AI governance frameworks, including the EU AI Act and NIST AI RMF, address safety, transparency, and accountability but do not operationalize

safetyarxiv-cs-ai
31 Jul 2026
Hardware

Together AI gives developers a high-throughput production path for running Inkling-Small on NVIDIA Accelerated Infrastructure across multimo…

DGX agent

Together AI gives developers a high-throughput production path for running Inkling-Small on NVIDIA Accelerated Infrastructure across multimodal, coding, and agentic workloads. Start building: https://

hardwaretogether-ai--x
31 Jul 2026
Model Releases

CMT-RAG: Complementary Memory Traces for Multi-turn Multi-hop RAG

DGX agent

arXiv:2607.26470v1 Announce Type: new Abstract: Multi-turn information-seeking conversations require both multi-hop reasoning and long-range dependency tracking across turns. However, existing RAG sys

model-releasesarxiv-cs-cl
30 Jul 2026
Applications

ContactFlow: A video action conditioning that transfers across embodiments

DGX agent

arXiv:2607.26579v1 Announce Type: cross Abstract: World models offer a promising route toward robot planning by enabling agents to imagine and verify the consequences of actions before execution. Howe

applicationsarxiv-cs-cv
30 Jul 2026
Safety

Forecasting Trajectory-Level Safety Risks in Black-Box Multi-Turn Interactions

DGX agent

arXiv:2607.26820v1 Announce Type: new Abstract: As large language models (LLMs) evolve from standalone assistants into autonomous agents, ensuring their safety requires shifting beyond pointwise risk

safetyarxiv-cs-lg
30 Jul 2026
Safety

Large-Scale ChatBot Validation Through Customer Digital Twin Simulations

DGX agent

arXiv:2607.26060v1 Announce Type: new Abstract: LLM-based chatbots are transforming customer service in regulated domains such as banking, but scalable and cost-effective validation remains a critical

safetyarxiv-cs-cl
30 Jul 2026
Safety

one of the top cybersecurity models, post-trained from open weights by @depthfirstlabs on @FireworksAI_HQ long-horizon RL is as much an infr…

DGX agent

one of the top cybersecurity models, post-trained from open weights by @depthfirstlabs on @FireworksAI_HQ long-horizon RL is as much an infra problem as a research one: 100+ turn rollouts, async/pipel

safetyfireworks-ai--x
30 Jul 2026
← Previous
1…283284285286287…374
Next →