AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlog
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,724 results
Safety

Domesticated, not feral: Why evolvable AI is not yet a Darwinian threat Writing in PNAS (Proceedings of the National Academy of Sciences), t…

DGX agent

Domesticated, not feral: Why evolvable AI is not yet a Darwinian threat Writing in PNAS (Proceedings of the National Academy of Sciences), three heavyweights in AI and evolutionary biology criticized

safetygary-marcus--x
21 Jul 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Generosity Under Conditions: Hardening Google Cloud Access Management

DGX agent

In Google Cloud, Identity and Access Management (IAM) helps you maintain access control over your cloud resources and operations. While it includes other features, this is its primary purpose. If you

model-releasesgoogle-cloud-ai
21 Jul 2026
Applications

Most startups celebrate their first couple million of revenue. @FactoryAI gave it back. They didn’t have to. They chose to. They decided tha…

DGX agent

Most startups celebrate their first couple million of revenue. @FactoryAI gave it back. They didn’t have to. They chose to. They decided that the product just wasn’t good enough yet, and they wanted t

applicationssonya-huang--x
21 Jul 2026
Model Releases

My OCR model mislabels section titles as body text. Is a CRF the right fix, or am I overcomplicating it? [P]

DGX agent

Hi everyone, I'm working on extracting the hierarchical structure of long PDF documents (legal/regulatory text, lots of numbered sections) and would like to gather some feedback on my approach before

model-releasesr-machinelearning
21 Jul 2026
Model Releases

Sakana AI is doubling down on a thesis we believe will play an increasingly important role in AI development: the future of AI won't be defi…

DGX agent

Sakana AI is doubling down on a thesis we believe will play an increasingly important role in AI development: the future of AI won't be defined by a single frontier model, but by how intelligently man

model-releasesdavid-ha--x
21 Jul 2026
Industry

As always, Cosmos 3 Edge is fully open. This includes model weights, post-training recipes and code. Available now on @huggingface 🤗 https:…

DGX agent

NVIDIA AI has released Cosmos 3 Edge, a 4‑billion‑parameter open‑source world model designed for on‑device inference. It enables robots to learn and act, autonomous vehicles to interpret road scenes a

industryclem-delangue--x
20 Jul 2026
Tutorials

// Global Workspace in LLMs // arXiv paper for the popular J-space work from Anthropic. (bookmark it) The short recap: If you build on chain…

DGX agent

// Global Workspace in LLMs // arXiv paper for the popular J-space work from Anthropic. (bookmark it) The short recap: If you build on chain-of-thought or steering vectors, this work provides a mechan

tutorialsdair-ai--x
20 Jul 2026
Model Releases

Open ecosystem or walled garden? For @anthropicai's @katelyn_lesse and @angjiang the answer is clear. 'We actually aren't precious about “Yo…

DGX agent

Open ecosystem or walled garden? For @anthropicai's @katelyn_lesse and @angjiang the answer is clear. 'We actually aren't precious about “You should run these things on our infrastructure.' In practic

model-releasessonya-huang--x
20 Jul 2026
Model Releases

People are wasting their AI subscriptions. They don’t realize how powerful AI can be in achieving their own goals. They’re asking AI to repl…

DGX agent

People are wasting their AI subscriptions. They don’t realize how powerful AI can be in achieving their own goals. They’re asking AI to reply to emails with zero goals in mind and nothing laddering up

model-releasesallie-k--miller--x
20 Jul 2026
Model Releases

Accuracy Without Grounding: Diagnosing Visual Dependency Dissociation in Video LLM Benchmarks

DGX agent

arXiv:2607.13305v1 Announce Type: cross Abstract: Benchmark accuracy in video large language models (LLMs) is often treated as evidence of visual understanding. We audit this assumption across twenty

model-releasesarxiv-cs-ai
16 Jul 2026
Safety

Audited Selective Verification for Risk-Controlled N-1 Thermal Contingency Screening under Deployment Shift

DGX agent

arXiv:2607.13221v1 Announce Type: cross Abstract: Real-time N-1 contingency screening in an energy management system trades assurance against cost: verifying every credible outage with full power flow

safetyarxiv-cs-ai
16 Jul 2026
Safety

Final Authority in AI Governance: Frontier-Provider Sovereignty and Action-Centered Deployer Governance

DGX agent

arXiv:2607.13040v1 Announce Type: cross Abstract: This paper examines where final authority should sit once capable AI systems are embedded in organizational workflows. It compares two governance mode

safetyarxiv-cs-ai
16 Jul 2026
Local Ai

GigaWorld-Policy-0.5: A Faster and Stronger WAM Empowered by AutoResearch

DGX agent

arXiv:2607.13960v1 Announce Type: new Abstract: World Action Models (WAMs) improve robot policy learning by jointly modeling actions and future visual observations, using future scene evolution as den

local-aiarxiv-cs-ro
16 Jul 2026
Model Releases

HRIBench: Benchmarking Interaction-Centric Human-Robot Collaboration

DGX agent

arXiv:2607.13056v1 Announce Type: cross Abstract: Current vision-language-action (VLA) benchmarks primarily evaluate isolated manipulation skills while leaving human-robot interaction structure largel

model-releasesarxiv-cs-lg
16 Jul 2026
Safety

LAPO: Leave-One-Turn Attribution for Self-Generated Process Rewards in Multi-Turn Search Reasoning

DGX agent

arXiv:2607.13501v1 Announce Type: new Abstract: Reinforcement learning for multi-turn search reasoning typically relies on terminal outcome rewards, which cannot distinguish useful, redundant, and har

safetyarxiv-cs-ai
16 Jul 2026
Model Releases

Not All Needles Are Found: How Fact Distribution and Don't Make It Up Prompts Shape Retrieval, Reasoning, and Hallucination in Long-Context LLMs

DGX agent

arXiv:2601.02023v2 Announce Type: replace-cross Abstract: As Large Language Models (LLMs) increasingly utilize massive context windows as working memory for autonomous tasks, their reliability fluctua

model-releasesarxiv-cs-ai
16 Jul 2026
Model Releases

NVIDIAとSakana AI、オープンモデルによるイノベーションのため協業拡大 本日、Sakana AIはNVIDIAとのコラボレーションを強化し、日本発の「集合知」の取り組みを次なるフェーズへ進めることを発表します。 私たちのマルチエージェント・オーケストレーションシステム…

DGX agent

NVIDIAとSakana AI、オープンモデルによるイノベーションのため協業拡大 本日、Sakana AIはNVIDIAとのコラボレーションを強化し、日本発の「集合知」の取り組みを次なるフェーズへ進めることを発表します。 私たちのマルチエージェント・オーケストレーションシステム「Sakana Fugu」に、Nemotronファミリーを含むNVIDIAのオープンモデル群を統合します。 Sakana

model-releasesdavid-ha--x
16 Jul 2026
Safety

Privacy Preserving Recommender Systems Balancing Personalization with Privacy

DGX agent

arXiv:2607.13328v1 Announce Type: cross Abstract: Personalized recommendation systems are central to modern e-commerce and retail platforms, but they typically rely on centralized storage of detailed

safetyarxiv-cs-ai
16 Jul 2026
Model Releases

Safeguard-Conditioned Uplift: Measuring Utility-Risk Frontiers for Dual-Use Biology Assistants

DGX agent

arXiv:2607.13039v1 Announce Type: cross Abstract: Safety evaluations for dual-use biology assistants often measure base-model capability, refusal behavior, or jailbreak success. These metrics miss a d

model-releasesarxiv-cs-ai
16 Jul 2026
Research

The Entanglement Wall: Activation-Space Probes as Risk Detectors, Not Context Adjudicators

DGX agent

arXiv:2607.13075v1 Announce Type: cross Abstract: Context can change whether a request is harmful without changing its topic or surface form. We ask whether residual-stream probes distinguish harmful

researcharxiv-cs-ai
16 Jul 2026
Model Releases

We’re excited to collaborate with NVIDIA to build the next generation of Fugu orchestration models together, by incorporating leading open-w…

DGX agent

We’re excited to collaborate with NVIDIA to build the next generation of Fugu orchestration models together, by incorporating leading open-weights models. Sakana AI Teams With NVIDIA to Advance Open M

model-releasesdavid-ha--x
16 Jul 2026
Model Releases

🥉 3rd place: CashFromChaos, by David Diaz (@davddiazm) CashFromChaos starts from a single seller input and automates everything up until a …

DGX agent

🥉 3rd place: CashFromChaos, by David Diaz (@davddiazm) CashFromChaos starts from a single seller input and automates everything up until a completed sale. You send a photo and a one-line clue, and Her

model-releasesnous-research--x
15 Jul 2026
Model Releases

An Empirical Study for Android-to-OpenHarmony GUI Test Migration

DGX agent

arXiv:2607.11245v2 Announce Type: replace-cross Abstract: To reduce the substantial engineering effort required to test the corresponding applications from Android to OpenHarmony, migrating existing G

model-releasesarxiv-cs-ai
15 Jul 2026
Model Releases

BREAKING: Grok 4.5 has climbed to #2 on the FrontierSWE benchmark. The result places Grok 4.5 among the world's top-performing AI models for…

DGX agent

BREAKING: Grok 4.5 has climbed to #2 on the FrontierSWE benchmark. The result places Grok 4.5 among the world's top-performing AI models for software engineering tasks, highlighting its growing streng

model-releaseselon-musk--x
15 Jul 2026
Model Releases

CoRe: A Comprehensive Framework for Cross-Image Comparative Reasoning in Vision-Language Models

DGX agent

arXiv:2607.12786v1 Announce Type: new Abstract: Cross-image comparative reasoning remains challenging for vision-language models (VLMs), especially when correct prediction requires fine-grained attrib

model-releasesarxiv-cs-cv
15 Jul 2026
Model Releases

Current efficient frontier of open models

DGX agent

Efficiency defined as score over active parameters. Removed all the models that were not on the pareto frontier. Yes I'm aware that artificialanalysis.ai aggregate benchmark isn't perfect, but I have

model-releasesr-localllama
15 Jul 2026
Model Releases

DeGuNet: Depth-Guided Ultra-Compact Backbones for Efficient LiDAR-Camera 3D Detection

DGX agent

arXiv:2607.12419v1 Announce Type: new Abstract: In autonomous driving perception, the fusion of LiDAR and camera modalities has become the dominant paradigm for 3D object detection. However, current m

model-releasesarxiv-cs-cv
15 Jul 2026
Safety

Directional Constraints for Efficient Exploration in Safe Reinforcement Learning

DGX agent

arXiv:2607.12784v1 Announce Type: cross Abstract: Reinforcement Learning has revolutionized the landscape of robotic research, allowing robust learning of complex robotic skills in simulation. However

safetyarxiv-cs-lg
15 Jul 2026
Model Releases

Dynamic Resource Allocation for Ensemble Determinization MCTS

DGX agent

arXiv:2607.13007v1 Announce Type: new Abstract: Simulation-based algorithms are especially suited for high-uncertainty environments such as adversarial board games with significant elements of randomn

model-releasesarxiv-cs-ai
15 Jul 2026
Model Releases

Egocentric Bias in Vision-Language Models

DGX agent

arXiv:2602.15892v2 Announce Type: replace-cross Abstract: Visual perspective taking--inferring how the world appears from another's viewpoint--is foundational to social cognition. We introduce FlipSet

model-releasesarxiv-cs-ai
15 Jul 2026
Model Releases

Environment Parameter Gradient Theorem for Policy-Environment Co-Design in Reinforcement Learning

DGX agent

arXiv:2607.12590v1 Announce Type: cross Abstract: Reinforcement learning (RL) is traditionally concerned with learning a control policy for a fixed environment. In many engineering systems, however, t

model-releasesarxiv-cs-lg
15 Jul 2026
Model Releases

Epistemic Stance Flexibility Probing: Measuring Prompt-Conditioned Register Shift in Large Language Models

DGX agent

arXiv:2607.12739v1 Announce Type: new Abstract: A language model may be asked either what experts believe about a contested claim or what it believes about the claim itself. A trustworthy conversation

model-releasesarxiv-cs-cl
15 Jul 2026
Research

EVOQUANT: Self-Evolving Verifier-Guided Strategy Optimization for Robust Quantitative Trading

DGX agent

arXiv:2607.12455v1 Announce Type: new Abstract: Quantitative strategy optimization remains largely manual, requiring domain experts to identify weak signals, tune risk-control rules, and repeatedly va

researcharxiv-cs-ai
15 Jul 2026
Safety

ExToken: Structured Exploration for Efficient Vision-Language-Action Reinforcement Fine-tuning

DGX agent

arXiv:2607.12931v1 Announce Type: new Abstract: Reinforcement Learning (RL) has demonstrated significant potential for improving Vision-Language-Action (VLA) models on complex manipulation tasks. Howe

safetyarxiv-cs-ro
15 Jul 2026
Model Releases

FairCoder: Probing LLM Bias in High-Stakes Decision Making via Coding Tasks

DGX agent

arXiv:2501.05396v3 Announce Type: replace Abstract: Large language models (LLMs) are increasingly used in high-stakes decisions such as hiring and college admissions, making their social bias a critic

model-releasesarxiv-cs-cl
15 Jul 2026
Model Releases

Good reason to try Grok 4.5 with Grok Build. It gets better every day!

DGX agent

Good reason to try Grok 4.5 with Grok Build. It gets better every day! Grok 4.5 just took the #1 spot on the Long-Horizon Terminal-Bench, outperforming Claude Fable 5, Claude Opus 4.8 and GPT-5.6-sol

model-releaseselon-musk--x
15 Jul 2026
Model Releases

I’ve found @angjiang and @katelyn_lesse of @AnthropicAI to be generous, transparent, and thoughtful when it comes to building an ecosystem, …

DGX agent

I’ve found @angjiang and @katelyn_lesse of @AnthropicAI to be generous, transparent, and thoughtful when it comes to building an ecosystem, not a walled garden. Listen and decide for yourself. Just so

model-releasessonya-huang--x
15 Jul 2026
Safety

Learning When to Trust in Contextual Social Bandits

DGX agent

arXiv:2603.13356v2 Announce Type: replace Abstract: Robust reinforcement learning typically assumes that feedback sources are either globally trustworthy or corrupted within a fixed global budget. We

safetyarxiv-cs-ai
15 Jul 2026
Research

LLMs Can See the Smoke but not the Fire: Evaluating Abductive Reasoning with Elenchos

DGX agent

arXiv:2607.12733v1 Announce Type: new Abstract: Large language models (LLMs) excel at pattern recognition and text generation, but their capacity for abductive inference - inferring latent hypotheses

researcharxiv-cs-ai
15 Jul 2026
Safety

Mistake gating leads to energy and memory efficient continual learning

DGX agent

arXiv:2604.14336v2 Announce Type: replace Abstract: Synaptic plasticity is metabolically expensive, yet animals continuously update their internal models without exhausting energy reserves. However, w

safetyarxiv-cs-ai
15 Jul 2026
Applications

Mobility-Aware Cache Framework for Scalable LLM-Based Human Mobility Simulation

DGX agent

arXiv:2602.16727v2 Announce Type: replace Abstract: Simulating large-scale human mobility is fundamental to understanding population movement patterns and supporting real-world geospatial applications

applicationsarxiv-cs-ai
15 Jul 2026
Local Ai

More Than Where You Are: Learning Semantics, Structure, and Geometry from Cross-View Localization

DGX agent

arXiv:2607.12429v1 Announce Type: new Abstract: Consistent cross-view understanding under extreme viewpoint changes is essential for spatial intelligence, as it enables models to recognize the same sc

local-aiarxiv-cs-cv
15 Jul 2026
Model Releases

NeSy-Route: A Neuro-Symbolic Benchmark for Constrained Route Planning in Remote Sensing

DGX agent

arXiv:2603.16307v2 Announce Type: replace Abstract: Remote sensing underpins crucial applications such as disaster relief and ecological field surveys, where systems must understand complex scenes and

model-releasesarxiv-cs-ai
15 Jul 2026
Local Ai

RL post-training on 14 Macs across 4 countries

DGX agent

Disclosure: I work at Pluralis Research, the lab that built this. Code is open, and I'm happy to answer questions. TL;DR: As far as we can tell, this is the first RL post-training run whose entire rol

local-air-localllama
15 Jul 2026
Research

Semantic-Edge Response Decoding of SAM3 for Zero-Shot Crack Segmentation

DGX agent

arXiv:2607.12292v1 Announce Type: new Abstract: Crack segmentation is essential for infrastructure inspection and structural health assessment, but existing high-performance methods typically require

researcharxiv-cs-cv
15 Jul 2026
Model Releases

🆕This Year In Claude https://www.youtube.com/watch?v=uU5Gv2h8-9g @simonw chats with @_catwu and @trq212 about the state of: - @claudeai Cod…

DGX agent

🆕This Year In Claude https://www.youtube.com/watch?v=uU5Gv2h8-9g @simonw chats with @_catwu and @trq212 about the state of: - @claudeai Code - Claude Fable - @anthropicai culture & product strategy -

model-releasesswyx--x
15 Jul 2026
Applications

Training against GPT‑Red makes GPT‑5.6 substantially more resilient. To measure this, we replayed some of GPT‑Red’s strongest attacks—none o…

DGX agent

Training against GPT‑Red makes GPT‑5.6 substantially more resilient. To measure this, we replayed some of GPT‑Red’s strongest attacks—none of which our models had seen during training. GPT‑5.6 Sol pro

applicationsopenai--x
15 Jul 2026
Model Releases

Transforming LLMs into Efficient Cross-Encoders via Knowledge Distillation for RAG Reranking

DGX agent

arXiv:2607.11933v1 Announce Type: new Abstract: Cross-encoders achieve high reranking accuracy in Retrieval-Augmented Generation (RAG) pipelines but impose quadratic inference costs that limit real-ti

model-releasesarxiv-cs-cl
15 Jul 2026
← Previous
1…334335336337338…370
Next →