AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlog
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,771 results
Applications

Legal AI startup Harvey reportedly raising 500M at 15.5B valuation

DGX agent

Harvey AI Corp., a provider of artificial intelligence software for attorneys, is reportedly seeking at least 500 million in new funding. The Information today cited sources as saying that the round c

applicationssiliconangle
7 Aug 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Moonlight & Mayhem (Raccoon Heist by Codex + GPT-5.6 Sol Ultra)

DGX agent

Moonlight & Mayhem (Raccoon Heist by Codex + GPT-5.6 Sol Ultra) On Wednesday I wrote about One-shotting a Raccoon Heist game using Claude Fable 5, where I had Claude Fable 5 build a full working game

model-releasessimon-willison
7 Aug 2026
Model Releases

Adversarial Attacks for Good: A Survey of Proactive Protection across the Visual Content Lifecycle

DGX agent

arXiv:2608.04314v1 Announce Type: cross Abstract: Once visual content enters an AI pipeline, its owner often retains little technical control over how it is used. Legal and regulatory remedies can add

model-releasesarxiv-cs-cv
6 Aug 2026
Model Releases

An AI model from Meta also hacked another company during testing

DGX agent

An AI model from Meta also hacked another company during testing Stop me if you've heard this one before: An AI model from the parent company of Facebook and Instagram hacked into another company’s sy

model-releasessimon-willison
6 Aug 2026
Model Releases

BrainBench: Benchmarking Large Language Models for Comprehensive EEG Understanding

DGX agent

arXiv:2608.04156v1 Announce Type: new Abstract: Electroencephalography (EEG) analysis extends beyond assigning predefined labels to recordings; it requires workflows connecting natural-language instru

model-releasesarxiv-cs-ai
6 Aug 2026
Research

For a very long time most high-performing AI models were end-to-end neural models; vector input -> vector output, with only ultra-thin symbo…

DGX agent

For a very long time most high-performing AI models were end-to-end neural models; vector input -> vector output, with only ultra-thin symbolic preprocessing and postprocessing layers (e.g. label deco

researchfrancois-chollet--x
6 Aug 2026
Safety

HCRide: Harmonizing Passenger Fairness and Driver Preference for Human-Centered Ride-Hailing

DGX agent

arXiv:2508.04811v2 Announce Type: replace Abstract: Order dispatch systems play a vital role in ride-hailing services, which directly influence operator revenue, driver profit, and passenger experienc

safetyarxiv-cs-lg
6 Aug 2026
Industry

Inevitable AI Group raises $6M to derail SaaS incumbents with more agile, AI-native software startups

DGX agent

Aleph, one of Israel’s top funds, is backing a new artificial intelligence-native venture studio with 6 million in pre-seed funding so it can build and launch dozens of new software startups by the en

industrysiliconangle
6 Aug 2026
Model Releases

K-EXAONE 2.0 Technical Report

DGX agent

arXiv:2608.04505v1 Announce Type: new Abstract: This technical report presents K-EXAONE 2.0, an open-weight multilingual foundation model developed by LG AI Research as a step in our effort toward glo

model-releasesarxiv-cs-cl
6 Aug 2026
Model Releases

Meta’s Muse Spark 1.1 hacked an external organization during cybersecurity test

DGX agent

A large language model developed by Meta Platforms Inc. hacked a third party organization during a cybersecurity evaluation. The Facebook parent disclosed the incident on Wednesday without specifying

model-releasessiliconangle
6 Aug 2026
Safety

NSF-HRPT: Neural Semantic Field meets Hierarchical Risk Perception Tree for Safety-Critical Scenario Assessment

DGX agent

arXiv:2608.04776v1 Announce Type: new Abstract: The ability to accurately assess and anticipate risks in safety-critical scenarios is crucial for autonomous driving systems. While existing research ha

safetyarxiv-cs-ai
6 Aug 2026
Model Releases

nvidias nemotron omni only loads its text half on a mac, so i wrote the vision and audio towers in mlx

DGX agent

nvidias nemotron omni is open weights and it sees, hears and reasons. theres already a 4bit mlx quant on hugging face but only the text backbone loads with standard mlx tooling. the model card says it

model-releasesr-localllama
6 Aug 2026
Model Releases

The Yokai Learning Environment: Tracking Beliefs Over Space and Time

DGX agent

arXiv:2508.12480v3 Announce Type: replace Abstract: The ability to cooperate with unknown partners is a central challenge in cooperative AI and widely studied in the form of zero-shot coordination (ZS

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Evaluating Counterfactual Sensitivity to Patient Information in Medication-Safety Reasoning

DGX agent

arXiv:2608.03028v1 Announce Type: new Abstract: Applying a valid medication-safety rule when its patient-specific conditions are not met can produce an incorrect decision. Existing medical evaluations

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

One-shotting a Raccoon Heist game using Claude Fable 5

DGX agent

Back in 2024 I tweeted screenshots of a game concept generated by GPT-3 and some concept 'art' created using DALL-E. Today, on the fourth anniversary of that tweet, I decided to see if Claude Fable 5

model-releasessimon-willison
5 Aug 2026
Local Ai

When and Where to Look: Adaptive Visual Evidence Scheduling for Efficient Long Video Understanding

DGX agent

arXiv:2608.03918v1 Announce Type: cross Abstract: Efficient long-video understanding requires vision--language models (VLMs) to reason over a small number of frames selected as sparse visual evidence.

local-aiarxiv-cs-ai
5 Aug 2026
Model Releases

Abduction Without a Body? Representational Grounding and the Abduction Loop for Scientific Hypothesis Generation

DGX agent

arXiv:2608.02505v1 Announce Type: cross Abstract: Can scientific abduction occur without continuous sensorimotor embodiment? Recent arguments in AI and philosophy of science hold that genuine hypothes

model-releasesarxiv-cs-cv
4 Aug 2026
Tutorials

Automated web insight extraction with Amazon Bedrock AgentCore

DGX agent

Extracting insights from dozens of websites by hand quickly becomes overwhelming. This post shows how to build an automated web insight extraction solution with Amazon Bedrock AgentCore Browser, Amazo

tutorialsaws-ml-blog
4 Aug 2026
Hardware

Choosing the right orchestration layer for your AI use cases

DGX agent

Compute scarcity is not the only struggle AI teams face. Optimal utilization is also key. Friction also emerges when they outgrow informal coordination methods such as shared spreadsheets, manual SSH

hardwarelambda-labs
4 Aug 2026
Local Ai

Device-First Feedback: Toward Mobile-Native LLM-Driven Neural Architecture Search

DGX agent

arXiv:2608.00078v1 Announce Type: new Abstract: Deploying convolutional neural networks generated by large language models (LLMs) on real mobile hardware requires more than GPU validation accuracy: IN

local-aiarxiv-cs-cv
4 Aug 2026
Safety

How Deutsche Bank unlocked agility with an API-ready ecosystem

DGX agent

When people think about digital transformation in banking, they often focus on the visible results: mobile apps and new digital services. But there's an invisible infrastructure making all these servi

safetygoogle-cloud-ai
4 Aug 2026
Safety

How Target is enhancing retail discovery and cutting database maintenance by 50% with Spanner Graph

DGX agent

In today’s retail environment, shoppers expect highly personalized product discovery experiences and conversational assistance that feels genuine, natural, and genuinely helpful. Today, successful pro

safetygoogle-cloud-ai
4 Aug 2026
Safety

IACM-RL: Intent-Aware Context Management and Reinforcement Learning for Complex Tool Invocation under Dynamic Intent Fluctuations

DGX agent

arXiv:2608.02110v1 Announce Type: new Abstract: Executing long-horizon tool invocations in real-world environments is severely challenged by dynamic user intent noise. Existing methods attempt robustn

safetyarxiv-cs-cl
4 Aug 2026
Model Releases

MixedComplementarityProblems.jl: A Fast, Batched, Open-Source Interior Point Solver for Mixed Complementarity Problems

DGX agent

arXiv:2608.00959v1 Announce Type: cross Abstract: Mixed complementarity problems (MCPs) arise as the first-order optimality conditions of nonlinear programs and noncooperative games, and provide a nat

model-releasesarxiv-cs-ro
4 Aug 2026
Safety

Practical Noise Modeling for SPAD Intensity Imaging

DGX agent

arXiv:2608.00489v1 Announce Type: new Abstract: Single-photon avalanche diode (SPAD) cameras are promising for low-light and high-dynamic-range intensity imaging, but their practical use is limited by

safetyarxiv-cs-cv
4 Aug 2026
Model Releases

SCHEDBench: A Benchmark for Evaluating LLM Constraint Faithfulness in Natural-Language Combinatorial Scheduling

DGX agent

arXiv:2608.00991v1 Announce Type: cross Abstract: This paper introduces SCHEDBench, a natural-language benchmark for evaluating combinatorial scheduling constraint faithfulness under surface-form vari

model-releasesarxiv-cs-cl
4 Aug 2026
Safety

Self-Improving Large Language Models via Progressive Experience Evolution

DGX agent

arXiv:2608.02139v1 Announce Type: new Abstract: Large language models (LLMs) capable of self-improvement require not only effective policy optimization, but also a principled mechanism for transformin

safetyarxiv-cs-cl
4 Aug 2026
Safety

Today, we’re launching Alpamayo 2 Super, our frontier open reasoning model for autonomous vehicles. Beyond seeing, Alpamayo understands and …

DGX agent

Today, we’re launching Alpamayo 2 Super, our frontier open reasoning model for autonomous vehicles. Beyond seeing, Alpamayo understands and reasons through the complex world - thinks before it acts. I

safetyelon-musk--x
4 Aug 2026
Model Releases

Toward Robust LLM-Based Judges: Taxonomic Bias Evaluation and Debiasing Optimization

DGX agent

arXiv:2603.08091v2 Announce Type: replace Abstract: Large language model (LLM)-based judges are widely adopted for automated evaluation and reward modeling, yet their judgments are often affected by j

model-releasesarxiv-cs-cl
4 Aug 2026
Research

8月より、Sakana AI @SakanaAILabs にApplied Research Engineer Internship として入社しました 🐟🐠🐡! 大学の夏季休業期間にフルタイム勤務予定です! LLMの研究開発と社会実装を頑張ります💪🏻

DGX agent

Horiyuki 'horiyuki42' joined Sakana AI Labs in August 2026 as an Applied Research Engineer Intern. He plans to work full‑time during the university summer recess while concentrating on large‑language‑

researchdavid-ha--x
3 Aug 2026
Model Releases

Benchmarks Are Not Validation: A System-Level View of Financial LLM Applications

DGX agent

arXiv:2607.28840v1 Announce Type: new Abstract: Large language models are increasingly deployed in financial applications that combine retrieval, proprietary data, tool use, orchestration logic, monit

model-releasesarxiv-cs-cl
3 Aug 2026
Model Releases

Detecting Experiential Intertextuality Across Migration Routes: Beyond Surface Similarity in French Narratives

DGX agent

arXiv:2607.29188v1 Announce Type: new Abstract: Migrants traversing geographically distinct routes such as the Trans-Saharan and Balkan corridors often recount strikingly parallel lived experiences: p

model-releasesarxiv-cs-cl
3 Aug 2026
Safety

Do Medical Foundation Models Generalize on the African Brain?

DGX agent

arXiv:2607.28771v1 Announce Type: new Abstract: Medical foundation models (FMs) are increasingly used for brain MRI analysis. However, their evaluation remains dominated by high-resource datasets, lea

safetyarxiv-cs-cv
3 Aug 2026
Safety

Fragility of Value under Imperfect Alignment

DGX agent

arXiv:2607.28881v1 Announce Type: new Abstract: As more responsibility is placed upon AI systems, it becomes increasingly important to guarantee that these systems are aligned with humanity. A common

safetyarxiv-cs-ai
3 Aug 2026
Safety

Gated Q-learning: Add Off-Policy Bias to Taste

DGX agent

arXiv:2607.28916v1 Announce Type: cross Abstract: Multistep credit assignment is critical for sample-efficient reinforcement learning, yet managing off-policy bias in Q-learning remains a fundamental

safetyarxiv-cs-ai
3 Aug 2026
Local Ai

I got tired of ad-filled mobile wrappers for Ollama, so I built PocketLLM Lite an open-source, offline Android client (Local GGUF, SKILL.md plugins, local RAG)

DGX agent

Hey, Like a lot of people here, I use local models via Ollama on my desktop/server and wanted a mobile client that actually felt responsive, worked offline, and respected privacy. Most apps on the Pla

local-air-ollama
3 Aug 2026
Model Releases

Identifying Informative Environments for Cognition Parameter Inference via Bayesian Experimental Design

DGX agent

arXiv:2607.28894v1 Announce Type: new Abstract: Computational cognitive modeling seeks to infer latent cognitive mechanisms underlying observed behavior. Bayesian inverse planning provides a principle

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

Linear Proposal Operators and Stochastic Search Geometry in SOMA and Differential Evolution

DGX agent

arXiv:2607.29228v1 Announce Type: cross Abstract: Swarm and evolutionary algorithms are usually analyzed as complete procedural systems in which nonlinear selection, replacement, and adaptation obscur

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

Looks Right, Works Right: A Project-Level Benchmark for Multi-Screen Mobile App Generation

DGX agent

arXiv:2607.28645v1 Announce Type: cross Abstract: Recent multimodal large language models can convert visual designs directly into executable code, but real mobile products require multiple screenshot

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

LWiAI Podcast #253 - Opus 5, Gemini 3.6, Kimi K3, Hugging Face Hack

DGX agent

LWiAI Podcast #253 (July 29 2026) covers a roundup of recent AI developments: Anthropic introduced Claude Opus 5 with Fable‑like capabilities; Google released Gemini 3.6/3.5 “Flash” variants and a cyb

model-releaseslast-week-in-ai
3 Aug 2026
Local Ai

Reflection or Re-Generation? Why LLM Revision Fails Where Human Revision Succeeds

DGX agent

arXiv:2607.28908v1 Announce Type: new Abstract: Reflection, the ability to revisit and revise prior reasoning, is central to how humans improve their answers. Large language models (LLMs) are increasi

local-aiarxiv-cs-lg
3 Aug 2026
Model Releases

TAPR: Enhancing LLM Performance with a Task-Aware Prompt Rewriter

DGX agent

arXiv:2607.28657v1 Announce Type: new Abstract: Large Language Models (LLMs) often require carefully crafted prompts to unlock their full potential, which can be a barrier for non-expert users. This w

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

To Add Is Machine, To Delete Is Human: Measuring and Mitigating Deletion Avoidance in LLM Code Editing

DGX agent

arXiv:2607.28887v1 Announce Type: cross Abstract: Large language models increasingly write and repair production code, yet evidence is mounting that their test-passing patches leave codebases harder t

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

two weeks ago i went on @swyx's pod and said some things that i... should not have said. a lot has happened since then, i owe you all an apo…

DGX agent

two weeks ago i went on @swyx's pod and said some things that i... should not have said. a lot has happened since then, i owe you all an apology. i'm sorry that i was right about every single thing. a

model-releasesswyx--x
3 Aug 2026
Model Releases

All other models on the Portal remain 20% discounted, aside from GPT-5.6 Terra and Luna which are 50% off. https://x.com/NousResearch/status…

DGX agent

All other models on the Portal remain 20% discounted, aside from GPT-5.6 Terra and Luna which are 50% off. https://x.com/NousResearch/status/2080039066771337475?s=20 All models are now 20% off for a l

model-releasesnous-research--x
2 Aug 2026
Model Releases

Running DeepSeek-V4-Flash-0731 (155 GB MoE) on a DGX Spark with vLLM-Moet 2-bit quantization - AI's narrative

DGX agent

# Running DeepSeek-V4-Flash-0731 (155 GB MoE) on a DGX Spark with vLLM-Moet 2-bit quantization I used Deepseek-v4-Flash-0731 cloud API settig up vllm-moet to run deepseek-v4-flash with MTP locally on

model-releasesr-localllama
2 Aug 2026
Research

To address the limits of deep learning and avoid stalling, the field of AI started by applying patch (1), which started being demoed 9 month…

DGX agent

To address the limits of deep learning and avoid stalling, the field of AI started by applying patch (1), which started being demoed 9 months later in December 2024 and has now become completely ubiqu

researchfrancois-chollet--x
2 Aug 2026
Safety

A Systems Engineering Framework for Vision-Language-Enabled UAV Triage and Disaster Response

DGX agent

arXiv:2607.27597v1 Announce Type: new Abstract: Recent advances in Vision Language Models (VLMs) have created new opportunities for disaster response, where responders must interpret large volumes of

safetyarxiv-cs-ro
31 Jul 2026
← Previous
1…302303304305306…371
Next →