AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,964 results
Agents

The most important thing about Grok Build and the 4.5 release is that it is genuinely so useful for real-world work

DGX agent

The most important thing about Grok Build and the 4.5 release is that it is genuinely so useful for real-world work Grok 4.5 just topped Perplexity’s WANDR orchestrator evaluation It scored higher tha

agentselon-musk--x
10 Jul 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Agents

When Does Continual Learning Require Learning

DGX agent

arXiv:2607.07847v1 Announce Type: new Abstract: As large language models (LLMs) become increasingly capable, the next question is how can we enable models to continually learn? Today, the field largel

agentsarxiv-cs-lg
10 Jul 2026
Agents

Gimitest: A Comprehensive Tool for Testing Reinforcement Learning Policies

DGX agent

arXiv:2607.07029v1 Announce Type: cross Abstract: Reinforcement learning (RL) policies can be unsafe and vulnerable to attacks. Ensuring their reliability is often a pain point as existing automated t

agentsarxiv-cs-ai
9 Jul 2026
Agents

Grok 4.5 on OpenClaw

DGX agent

Grok 4.5 on OpenClaw Grok 4.5 from @SpaceXAI is live on OpenClaw. No OpenClaw update required, just connect your X Premium or SuperGrok subscription, select Grok 4.5 under the xAI provider, and use an

agentselon-musk--x
9 Jul 2026
Agents

Measuring Intelligence Beyond Human Scale

DGX agent

arXiv:2607.07040v1 Announce Type: new Abstract: How can we measure intelligence beyond human capability? Human-authored benchmarks saturate, and above human capability, examiners may not know which ta

agentsarxiv-cs-ai
9 Jul 2026
Model Releases

My feed has been hijacked this week by frontier model influencers who say they've had 5.6 Sol and Fable for 'months'. I think this tells an …

DGX agent

My feed has been hijacked this week by frontier model influencers who say they've had 5.6 Sol and Fable for 'months'. I think this tells an inaccurate story the field. From the outside it makes AI pro

model-releasesharrison-chase--x
9 Jul 2026
Agents

Neutral Substrates: A Design Constraint for Shared Records Under Persistent Interpretive Disagreement

DGX agent

arXiv:2601.14271v2 Announce Type: replace Abstract: Shared accountability records are often used by parties who may never agree about causation, responsibility, or normative interpretation. For such r

agentsarxiv-cs-ai
9 Jul 2026
Agents

Recursive Language Models Meet Uncertainty: The Surprising Effectiveness of Self-Reflective Program Search for Long Context

DGX agent

Long-context handling remains a core challenge for language models: even with extended context windows, models often fail to reliably extract, reason over, and use the information across long contexts

agentsapple-ml-research
9 Jul 2026
Agents

Thank you @hwchase17 @BraceSproul @devstein64 @jeffreyhuber for the LLM Wikis session today, I absolutely loved it! Really grateful you keep…

DGX agent

Thank you @hwchase17 @BraceSproul @devstein64 @jeffreyhuber for the LLM Wikis session today, I absolutely loved it! Really grateful you keep these sessions open and honest about what actually works in

agentsharrison-chase--x
9 Jul 2026
Agents

Token per watt becomes the defining metric as storage moves to AI’s critical path

DGX agent

Token per watt — not raw compute — is emerging as the defining efficiency metric for AI data centers, putting storage at the center of an infrastructure rethink that is reshaping how the industry meas

agentssiliconangle
9 Jul 2026
Agents

EvalLoop: A Methodology for Evaluation-Driven Iterative Improvement of Business AI Systems

DGX agent

arXiv:2607.05638v1 Announce Type: cross Abstract: Teams deploying large language models in business contexts need evaluation systems, yet most treat evaluation as static model selection: run benchmark

agentsarxiv-cs-ai
8 Jul 2026
Model Releases

Evaluating calibrated refusal and safe usefulness in dual-use biology settings

DGX agent

arXiv:2607.05462v1 Announce Type: cross Abstract: As AI agents are incorporated into life science workflows, the capabilities that speed discovery might also enable misuse. We present BioSecBench-Refu

model-releasesarxiv-cs-ai
8 Jul 2026
Agents

Faithful or Findable? Evaluating LLM-Generated Metadata for RDF Dataset Search

DGX agent

arXiv:2607.05970v1 Announce Type: cross Abstract: Dataset search depends heavily on metadata, making LLM-generated metadata a consequential form of synthetic content in retrieval systems. We study six

agentsarxiv-cs-ai
8 Jul 2026
Safety

From Passive Retrieval to Active Memory Navigation: Learning to Use Memory as a Structured Action Space

DGX agent

arXiv:2607.05794v1 Announce Type: new Abstract: Long-term user memory is essential for personalized conversational agents, yet many memory systems still expose memory through passive retrieval interfa

safetyarxiv-cs-ai
8 Jul 2026
Agents

Learning The Minimum Action Distance

DGX agent

arXiv:2506.09276v4 Announce Type: replace-cross Abstract: This paper presents a state representation framework for Markov decision processes (MDPs) that can be learned solely from state trajectories,

agentsarxiv-cs-ai
8 Jul 2026
Model Releases

Quantifying Frontier LLM Capabilities for Container Sandbox Escape

DGX agent

arXiv:2603.02277v2 Announce Type: replace-cross Abstract: Large language models (LLMs) increasingly act as autonomous agents, using tools to execute code, read and write files, and access networks, cr

model-releasesarxiv-cs-ai
8 Jul 2026
Model Releases

Rewriting Bun in Rust

DGX agent

Rewriting Bun in Rust Jarred Sumner has been promising this blog post (since May 9th) about his Zig to Rust rewrite of Bun for significantly longer than it took him to finish the rewrite. Honestly, it

model-releasessimon-willison
8 Jul 2026
Agents

SecureCode: A Production-Grade Multi-Turn Dataset for Training Security-Aware Code Generation Models

DGX agent

arXiv:2512.18542v3 Announce Type: replace-cross Abstract: AI coding assistants produce vulnerable code in 45% of security-relevant scenarios~ite{veracode2025}, yet no public training dataset teaches b

agentsarxiv-cs-ai
8 Jul 2026
Agents

VaseMuseum: Digital Intelligent Museum for Ancient Greek Pottery

DGX agent

arXiv:2607.06374v1 Announce Type: new Abstract: Vision-language models (VLMs) have made interactive digital museums increasingly feasible by connecting 3D digitization with natural-language artifact e

agentsarxiv-cs-cv
8 Jul 2026
Agents

You don't need heavyweight VLMs to OCR simple text-only PDFs. Doing that is like bringing a bazooka to a knife-fight, and is completely unne…

DGX agent

You don't need heavyweight VLMs to OCR simple text-only PDFs. Doing that is like bringing a bazooka to a knife-fight, and is completely unnecessary and worse quality than a tuned OCR approach. Output

agentsjerry-liu--x
8 Jul 2026
Safety

A Mathematical Theory of Value: a synthesis on goal-directed agency under resource constraints

DGX agent

arXiv:2606.12502v2 Announce Type: replace-cross Abstract: We propose that value -- the quantity goal-directed agents create, destroy, and exchange -- is a lawful structural quantity in the same catego

safetyarxiv-cs-ai
7 Jul 2026
Agents

Conflict-Based Lazy Search for Fast Multi-Manipulator Planning

DGX agent

arXiv:2607.04124v1 Announce Type: cross Abstract: Employing multiple manipulators can boost efficiency and accomplish tasks that a single manipulator cannot do. However, real-time planning for multipl

agentsarxiv-cs-ai
7 Jul 2026
Model Releases

Determinants and Limits of LLM Security-Tool Orchestration: A Study with HexStrike-AI

DGX agent

arXiv:2607.02873v1 Announce Type: cross Abstract: Large language model agents driving security tool suites over the Model Context Protocol are increasingly common. Yet the factors that bound their cap

model-releasesarxiv-cs-ai
7 Jul 2026
Agents

Flow-A11y: Flow-Aware Accessibility Testing

DGX agent

arXiv:2607.03100v1 Announce Type: cross Abstract: Modern web applications increasingly expose accessibility barriers through interaction flows rather than static page snapshots. Keyboard traps, focus

agentsarxiv-cs-ai
7 Jul 2026
Agents

Governed Caste Reassignment in Heterogeneous Swarms: An Asymmetric-Trust Protocol with Audited Operator Countersignature

DGX agent

arXiv:2607.04634v1 Announce Type: new Abstract: In heterogeneous robot swarms, caste reassignment (rebinding a robot to a new capability-bound role) is a high-frequency runtime event driven by battery

agentsarxiv-cs-ro
7 Jul 2026
Agents

Heaviside Continuity of Rolling Coefficients for Eliminating Epistemic Entropy in Large Language Models

DGX agent

arXiv:2607.04562v1 Announce Type: new Abstract: Large language models (LLMs) generate fluent outputs that can be wrong. Unlike humans, who often exhibit cues when providing false information, LLMs pro

agentsarxiv-cs-ai
7 Jul 2026
Local Ai

Multi-Large Language Model Orchestrated Severity Assessment of Clinical Records (MOSAIC)

DGX agent

arXiv:2607.05032v1 Announce Type: new Abstract: Background: Disease severity is a multidimensional construct difficult to capture with rule-based approaches in Electronic Healthcare Records (EHR). Age

local-aiarxiv-cs-cl
7 Jul 2026
Agents

PromptPET: Privacy-Utility Optimized Prompt Obfuscation

DGX agent

arXiv:2607.02932v1 Announce Type: cross Abstract: Privacy is an important challenge when users interact with AI chatbots, since users may share sensitive information, explicitly or implicitly, and AI

agentsarxiv-cs-ai
7 Jul 2026
Agents

SMOCS: A Streaming Framework for Simplified Deployment, Monitoring, and Optimization of ML Systems in Production

DGX agent

arXiv:2607.02731v1 Announce Type: cross Abstract: Machine learning has demonstrated significant potential for real-time monitoring, optimization, and control of scientific facilities. However, deployi

agentsarxiv-cs-ai
7 Jul 2026
Agents

TGRIP: A Text-Guided Approach to Vehicle Instance Prediction in Autonomous Driving

DGX agent

arXiv:2607.04812v1 Announce Type: new Abstract: Bird's-Eye View (BEV) end-to-end instance prediction has emerged as a robust paradigm for autonomous driving perception, effectively mitigating the erro

agentsarxiv-cs-cv
7 Jul 2026
Agents

This is the great thing about OpenWiki! It’ll maintain the docs for you automatically. Just setup a GitHub to run: openwiki —update and it’l…

DGX agent

This is the great thing about OpenWiki! It’ll maintain the docs for you automatically. Just setup a GitHub to run: openwiki —update and it’ll put up a PR once a day with changes / additions to your me

agentsharrison-chase--x
7 Jul 2026
Model Releases

When Aggregate Alignment Misleads: Auditing Policy Repair Without Per-State Expert Actions

DGX agent

arXiv:2607.03386v1 Announce Type: new Abstract: Agentic AI systems are increasingly used to edit, refine, and repair decision policies, but evaluating these edits is difficult when per-state expert ac

model-releasesarxiv-cs-ai
7 Jul 2026
Agents

big things happened over the weekend!

DGX agent

big things happened over the weekend! OpenWiki is at 1.7k stars in just 3 days! Right now it's just for codebases, but we're working to expand it to everything for memory. What do you want to see in a

agentsharrison-chase--x
6 Jul 2026
Agents

How do you trace one number in a 200 page ESG report back to the exact page it came from? We dug into that with the @llama_index team behind…

DGX agent

How do you trace one number in a 200 page ESG report back to the exact page it came from? We dug into that with the @llama_index team behind LiteParse. We tested five ways to retrieve evidence across

agentsjerry-liu--x
6 Jul 2026
Agents

I predict 50% of companies will need new leadership, because the old management style won't work in the era of AI. That's exactly why I'm se…

DGX agent

I predict 50% of companies will need new leadership, because the old management style won't work in the era of AI. That's exactly why I'm seeing 95%+ of so-called 'AI Transformation' initiatives fail.

agentskai-fu-lee--x
6 Jul 2026
Agents

Run MiniMax models on Amazon Bedrock

DGX agent

In this post, we walk through how to get started with MiniMax models on Amazon Bedrock, including the capabilities supported by these models, the service tiers available, how on-demand inference scale

agentsaws-ml-blog
6 Jul 2026
Agents

.@aiDotEngineer World's Fair was one of the most unique, interesting conferences I've been to: - incredible conversations with builders - hi…

DGX agent

.@aiDotEngineer World's Fair was one of the most unique, interesting conferences I've been to: - incredible conversations with builders - hilarious & creative touches (shout out @swyx) from a flash mo

agentsjerry-liu--x
3 Jul 2026
Agents

Autonomous discovery of traffic laws with AI traffic scientists

DGX agent

arXiv:2607.01639v1 Announce Type: new Abstract: Universal traffic laws describe recurrent patterns in congestion, mobility and driving behavior across cities, providing a scientific basis for transpor

agentsarxiv-cs-ai
3 Jul 2026
Agents

Beyond Supervised Clarification: Input Rewriting with LLMs for Dialogue Discourse Parsing

DGX agent

arXiv:2607.01964v1 Announce Type: new Abstract: Rewriting inputs to improve frozen downstream models has become a common strategy in modern NLP pipelines. Prior work on incremental dialogue discourse

agentsarxiv-cs-cl
3 Jul 2026
Agents

Certified World Models as Sensing Clocks: Drift-Aware Deadlines for Active Perception

DGX agent

arXiv:2607.01537v1 Announce Type: new Abstract: Certified world models estimate how long their predictions remain valid. We turn this validity horizon into an operational sensing clock: a rule for whe

agentsarxiv-cs-lg
3 Jul 2026
Agents

CommonRoad-Game: A Human-in-the-Loop Simulation Framework for Autonomous Driving

DGX agent

arXiv:2607.01382v1 Announce Type: new Abstract: Motion planning algorithms should be evaluated in human-in-the-loop environments to ensure they produce safe and efficient behaviors during interactions

agentsarxiv-cs-ro
3 Jul 2026
Model Releases

ContextSniper: AntTrail's Token-Efficient Code Memory for Repository-Level Program Repair

DGX agent

arXiv:2607.01916v1 Announce Type: new Abstract: Large language model agents can repair real repository issues, but they often spend large context budgets on whole-file reads, broad searches, and long

model-releasesarxiv-cs-ai
3 Jul 2026
Agents

Grounded Optimization: A Layered Engineering Framework for Reducing LLM Hallucination in Automated Personal Document Rewriting

DGX agent

arXiv:2607.01457v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly applied to resume optimization for applicant tracking systems, introducing hallucination failures distin

agentsarxiv-cs-ai
3 Jul 2026
Agents

Language Models as Measurement Apparatus for Culture

DGX agent

arXiv:2607.02459v1 Announce Type: new Abstract: Language models are increasingly used to quantify cultural phenomena, but what makes such measurement distinctively cultural? This paper argues that NLP

agentsarxiv-cs-cl
3 Jul 2026
Agents

Ophiuchus: Incentivizing Tool-augmented 'Think with Images' for Joint Medical Segmentation, Understanding and Reasoning

DGX agent

arXiv:2512.14157v2 Announce Type: replace Abstract: Recent medical MLLMs have made significant progress in generating step-by-step textual reasoning chains. However, they still struggle with complex c

agentsarxiv-cs-ai
3 Jul 2026
Model Releases

OPINE-World: Programmatic World Modeling with Ontology-error-Prioritized Interactive Exploration

DGX agent

arXiv:2607.01531v1 Announce Type: new Abstract: Learning how an environment behaves from interaction is central to building agents that adapt to unfamiliar tasks. World models learned with deep networ

model-releasesarxiv-cs-ai
3 Jul 2026
Agents

Prompt engineering is costing you money. Learn how to fine-tune your models. Stuffing a lot of text into every single API call slows down yo…

DGX agent

Prompt engineering is costing you money. Learn how to fine-tune your models. Stuffing a lot of text into every single API call slows down your app because the model has to process all those tokens bef

agentsfireworks-ai--x
3 Jul 2026
Agents

Recursive Models for Long-Horizon Reasoning

DGX agent

arXiv:2603.02112v2 Announce Type: replace-cross Abstract: Modern language models reason within bounded context, an inherent constraint that poses a fundamental barrier to long-horizon reasoning. We id

agentsarxiv-cs-cl
3 Jul 2026
← Previous
1…230231232233234…375
Next →