AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
25,429 results
16 Apr 2026

HOLY 🤯 The one and only @elder_plinius just dropped an unlocked Gemma 4 E4B, and the specs are INSANE. Look at the performance shifts: → Re…

Model ReleasesDGX agent

HOLY 🤯 The one and only @elder_plinius just dropped an unlocked Gemma 4 E4B, and the specs are INSANE. Look at the performance shifts: → Refusal rate: 98.8% down to 2.1% (!!) → Compliance: 1.2% up to

How WPP accelerates humanoid robot training 10x with G4 VMs

Model ReleasesDGX agent

Editor’s note: Today we hear from Perry Nightingale, SVP of Creative AI at WPP about the workflow that cuts training time for humanoid robots from days to minutes — plus access to the open-source code

I've got a lot to say here! I'll post a guide on how to prompt with Opus 4.7 as well as talk about a personal project I've been working on u…

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Tutorials
DGX agent

I've got a lot to say here! I'll post a guide on how to prompt with Opus 4.7 as well as talk about a personal project I've been working on using it. I hope you enjoy working with Opus 4.7 and getting

LaoBench: A Large-Scale Multidimensional Lao Benchmark for Large Language Models

Model ReleasesDGX agent

arXiv:2511.11334v3 Announce Type: replace Abstract: The rapid advancement of large language models (LLMs) has not been matched by their evaluation in low-resource languages, especially Southeast Asian

LEO-RobotAgent: A General-purpose Robotic Agent for Language-driven Embodied Operator

AgentsDGX agent

arXiv:2512.10605v2 Announce Type: replace Abstract: We propose LEO-RobotAgent, a general-purpose language-driven intelligent agent framework for robots. Under this framework, LLMs can operate differen

Lite Any Stereo: Efficient Zero-Shot Stereo Matching

ApplicationsDGX agent

arXiv:2511.16555v3 Announce Type: replace Abstract: Recent advances in stereo matching have focused on accuracy, often at the cost of significantly increased model size. Traditionally, the community h

mcpstrike – Let your local LLM run the pentest for you

Local AiDGX agent

**mcpstrike** is a tool shared on the r/ollama subreddit that integrates a locally-run LLM (via Ollama) with an MCP (Model Context Protocol) server to autonomously execute penetration testing workflow

MERRIN: A Benchmark for Multimodal Evidence Retrieval and Reasoning in Noisy Web Environments

Model ReleasesDGX agent

arXiv:2604.13418v1 Announce Type: new Abstract: Motivated by the underspecified, multi-hop nature of search queries and the multimodal, heterogeneous, and often conflicting nature of real-world web re

Mozilla launches Thunderbolt AI client with focus on self-hosted infrastructure

Model ReleasesDGX agent

Thunderbolt is a new open-source AI client from Mozilla-owned MZLA Technologies aimed at enterprises who want to run self-hosted chatbots on their own infrastructure. The platform allows organizations

MulDimIF: A Multi-Dimensional Constraint Framework for Evaluating and Improving Instruction Following in Large Language Models

Model ReleasesDGX agent

arXiv:2505.07591v2 Announce Type: replace Abstract: Instruction following refers to the ability of large language models (LLMs) to generate outputs that satisfy all specified constraints. Existing res

Multi-Task LLM with LoRA Fine-Tuning for Automated Cancer Staging and Biomarker Extraction

Model ReleasesDGX agent

arXiv:2604.13328v1 Announce Type: new Abstract: Pathology reports serve as the definitive record for breast cancer staging, yet their unstructured format impedes large-scale data curation. While Large

OpenAI starts offering a biology-tuned LLM

IndustryDGX agent

OpenAI has launched GPT-Rosalind, a biology-tuned large language model . This specialized LLM is designed to enhance performance on biology-specific tasks and applications. The model represents OpenAI

OpenAI updates its Codex desktop app with features like computer control, an in-app browser, image generation, automation memory, plugin support, and more (David Gewirtz/ZDNET)

IndustryDGX agent

David Gewirtz / ZDNET: OpenAI updates its Codex desktop app with features like computer control, an in-app browser, image generation, automation memory, plugin support, and more — ZDNET's key takeaway

OPTED: Open Preprocessed Trachoma Eye Dataset Using Zero-Shot SAM 3 Segmentation

Model ReleasesDGX agent

arXiv:2603.06885v2 Announce Type: replace Abstract: Trachoma remains the leading infectious cause of blindness worldwide, with Sub-Saharan Africa bearing over 85% of the global burden and Ethiopia alo

Out of Context: Reliability in Multimodal Anomaly Detection Requires Contextual Inference

Model ReleasesDGX agent

arXiv:2604.13252v1 Announce Type: new Abstract: Anomaly detection aims to identify observations that deviate from expected behavior. Because anomalous events are inherently sparse, most frameworks are

Remember Me, Refine Me: A Dynamic Procedural Memory Framework for Experience-Driven Agent Evolution

AgentsDGX agent

arXiv:2512.10696v2 Announce Type: replace-cross Abstract: Procedural memory enables large language model (LLM) agents to internalize 'how-to' knowledge, theoretically reducing redundant trial-and-erro

Representation over Routing: Overcoming Surrogate Hacking in Multi-Timescale PPO

SafetyDGX agent

arXiv:2604.13517v1 Announce Type: new Abstract: Temporal credit assignment in reinforcement learning has long been a central challenge. Inspired by the multi-timescale encoding of the dopamine system

RFK Jr. forces FDA to reconsider 12 unproven peptides after 2023 ban

SafetyDGX agent

In 2023, the FDA removed 19 peptides from the list of drugs that compounding pharmacies could produce, and in 2026 the FDA announced it will review whether to add back 7 of these peptides following pr

Soft Q(lambda): A multi-step off-policy method for entropy regularised reinforcement learning using eligibility traces

SafetyDGX agent

arXiv:2604.13780v1 Announce Type: new Abstract: Soft Q-learning has emerged as a versatile model-free method for entropy-regularised reinforcement learning, optimising for returns augmented with a pen

Spectral methods: crucial for machine learning, natural for quantum computers?

SafetyDGX agent

arXiv:2603.24654v2 Announce Type: replace-cross Abstract: This article presents an argument for why quantum computers could unlock new methods for machine learning. We argue that spectral methods, in

this is fun. been trying eeg headsets for 10 yrs, the hat makes so much sense :)

AgentsDGX agent

this is fun. been trying eeg headsets for 10 yrs, the hat makes so much sense :) you can now control things with your brain. literally. we're building the most wearable BCI on the planet, with @sabica

TSMC CEO C.C. Wei says TSMC checked with customers about AI demand and was reassured that it was still strong amid the Iran war, as it raises revenue forecasts (Wall Street Journal)

IndustryDGX agent

Wall Street Journal: TSMC CEO C.C. Wei says TSMC checked with customers about AI demand and was reassured that it was still strong amid the Iran war, as it raises revenue forecasts — Taiwan company ex

UMI-3D: Extending Universal Manipulation Interface from Vision-Limited to 3D Spatial Perception

SafetyDGX agent

arXiv:2604.14089v1 Announce Type: new Abstract: We present UMI-3D, a multimodal extension of the Universal Manipulation Interface (UMI) for robust and scalable data collection in embodied manipulation

Why dynamically routing multi-timescale advantages in PPO causes policy collapse (and a simple decoupled fix) [R]

SafetyDGX agent

This Reddit post discusses a known instability in PPO when advantage estimates operating across different temporal scales (e.g., short-horizon and long-horizon returns) are dynamically routed or mixed

15 Apr 2026

390M+ embeddings. 100K+ namespaces. Sustained P50 latency of ~60ms at ~40 QPS. 📈 @ZoomInfo used our Dedicated Read Nodes, now in GA, to byp…

ApplicationsDGX agent

390M+ embeddings. 100K+ namespaces. Sustained P50 latency of ~60ms at ~40 QPS. 📈 @ZoomInfo used our Dedicated Read Nodes, now in GA, to bypass the 'infrastructure wall' and ship real-time AI recommend

A Comparison of Reinforcement Learning and Optimal Control Methods for Path Planning

SafetyDGX agent

arXiv:2604.12628v1 Announce Type: cross Abstract: Path-planning for autonomous vehicles in threat-laden environments is a fundamental challenge. While traditional optimal control methods can find idea

A longitudinal health agent framework

SafetyDGX agent

arXiv:2604.12019v1 Announce Type: new Abstract: Although artificial intelligence (AI) agents are increasingly proposed to support potentially longitudinal health tasks, such as symptom management, beh

Agentic Discovery with Active Hypothesis Exploration for Visual Recognition

AgentsDGX agent

arXiv:2604.12999v1 Announce Type: new Abstract: We introduce HypoExplore, an agentic framework that formulates neural architecture discovery for visual recognition as a hypothesis-driven scientific in

@agi_inc Rsvp to be invited to our future events: https://luma.com/9zbd9pqq

AgentsDGX agent

AGI Inc. is promoting upcoming events and inviting interested individuals to RSVP through a Luma event page to secure invitations. Luma is a platform commonly used for organizing and managing event re

AISafetyBenchExplorer: A Metric-Aware Catalogue of AI Safety Benchmarks Reveals Fragmented Measurement and Weak Benchmark Governance

Model ReleasesDGX agent

arXiv:2604.12875v1 Announce Type: new Abstract: The rapid expansion of large language model (LLM) safety evaluation has produced a substantial benchmark ecosystem, but not a correspondingly coherent m

AMD, Arm, and Qualcomm invest 60M in London-based self-driving tech startup Wayve as part of an extension to its 1.2B Series D, announced in February (Kirsten Korosec/TechCrunch)

IndustryDGX agent

Kirsten Korosec / TechCrunch: AMD, Arm, and Qualcomm invest 60M in London-based self-driving tech startup Wayve as part of an extension to its 1.2B Series D, announced in February — Chipmakers AMD, Ar

Artificial Intelligence for Modeling and Simulation of Mixed Automated and Human Traffic

AgentsDGX agent

arXiv:2604.12857v1 Announce Type: new Abstract: Autonomous vehicles (AVs) are now operating on public roads, which makes their testing and validation more critical than ever. Simulation offers a safe

BarbieGait: An Identity-Consistent Synthetic Human Dataset with Versatile Cloth-Changing for Gait Recognition

ApplicationsDGX agent

arXiv:2604.12221v1 Announce Type: new Abstract: Gait recognition, as a reliable biometric technology, has seen rapid development in recent years while it faces significant challenges caused by diverse

Benchmarking Foundation Models with Retrieval-Augmented Generation in Olympic-Level Physics Problem Solving

Model ReleasesDGX agent

arXiv:2510.00919v3 Announce Type: replace-cross Abstract: Retrieval-augmented generation (RAG) with foundation models has achieved strong performance across diverse tasks, but their capacity for exper

Bro discuss ChatGPT birth 7 year ago.

IndustryDGX agent

This Reddit post from r/ChatGPT appears to be a community discussion reflecting on ChatGPT's origins, likely referencing the fact that OpenAI's foundational work began years before the public launch.

Characterizing Human Semantic Navigation in Concept Production as Trajectories in Embedding Space

ApplicationsDGX agent

arXiv:2602.05971v2 Announce Type: replace Abstract: Semantic representations can be framed as a structured, dynamic knowledge space through which humans navigate to retrieve and manipulate meaning. To

ContextLens: Modeling Imperfect Privacy and Safety Context for Legal Compliance

SafetyDGX agent

arXiv:2604.12308v1 Announce Type: new Abstract: Individuals' concerns about data privacy and AI safety are highly contextualized and extend beyond sensitive patterns. Addressing these issues requires

Continuous Knowledge Metabolism: Generating Scientific Hypotheses from Evolving Literature

SafetyDGX agent

arXiv:2604.12243v1 Announce Type: cross Abstract: Scientific hypothesis generation requires tracking how knowledge evolves, not just what is currently known. We introduce Continuous Knowledge Metaboli

Dataset Safety in Autonomous Driving: Requirements, Risks, and Assurance

SafetyDGX agent

arXiv:2511.08439v2 Announce Type: replace Abstract: Dataset integrity is fundamental to the safety and reliability of AI systems, especially in autonomous driving. This paper presents a structured fra

Empirical Evaluation of PDF Parsing and Chunking for Financial Question Answering with RAG

Model ReleasesDGX agent

arXiv:2604.12047v1 Announce Type: new Abstract: PDF files are primarily intended for human reading rather than automated processing. In addition, the heterogeneous content of PDFs, such as text, table

Every Picture Tells a Dangerous Story: Memory-Augmented Multi-Agent Jailbreak Attacks on VLMs

SafetyDGX agent

arXiv:2604.12616v1 Announce Type: new Abstract: The rapid evolution of Vision-Language Models (VLMs) has catalyzed unprecedented capabilities in artificial intelligence; however, this continuous modal

Fine-Tuning LLMs for Report Summarization: Analysis on Supervised and Unsupervised Data

HardwareDGX agent

arXiv:2503.10676v2 Announce Type: replace-cross Abstract: We study the efficacy of fine-tuning Large Language Models (LLMs) for the specific task of report (government archives, news, intelligence rep

From Plan to Action: How Well Do Agents Follow the Plan?

Model ReleasesDGX agent

arXiv:2604.12147v1 Announce Type: cross Abstract: Agents aspire to eliminate the need for task-specific prompt crafting through autonomous reason-act-observe loops. Still, they are commonly instructed

Generative Anonymization in Event Streams

Model ReleasesDGX agent

arXiv:2604.12803v1 Announce Type: new Abstract: Neuromorphic vision sensors offer low latency and high dynamic range, but their deployment in public spaces raises severe data protection concerns. Rece

Generative Refinement Networks for Visual Synthesis

Model ReleasesDGX agent

arXiv:2604.13030v1 Announce Type: new Abstract: While diffusion models dominate the field of visual generation, they are computationally inefficient, applying a uniform computational effort regardless

GigaCheck: Detecting LLM-generated Content via Object-Centric Span Localization

Local AiDGX agent

arXiv:2410.23728v3 Announce Type: replace Abstract: With the increasing quality and spread of LLM assistants, the amount of generated content is growing rapidly. In many cases and tasks, such texts ar

Global smartphone shipments fell 4.1% YoY in Q1, the first decline since 2023, amid a memory chip crunch; Samsung's shipments grew 3.6% and Apple's grew 3.3% (IDC)

IndustryDGX agent

IDC: Global smartphone shipments fell 4.1% YoY in Q1, the first decline since 2023, amid a memory chip crunch; Samsung's shipments grew 3.6% and Apple's grew 3.3% — Limited memory supply and record hi

Habitat Classification from Ground-Level Imagery Using Deep Neural Networks

Local AiDGX agent

arXiv:2507.04017v3 Announce Type: replace Abstract: Habitat assessment at local scales -- critical for enhancing biodiversity and guiding conservation priorities -- often relies on expert field survey

Hybrid quantum-classical computing gains momentum in Europe

IndustryDGX agent

The European Union is taking ambitious steps to coordinate efforts by its member countries to achieve global leadership in quantum computing. The EU Quantum Computing Act, which is scheduled to take e

I spent some time trying to distill all the complex factors impacting open models -- economics, capabilities, distribution, policy, etc. -- …

SafetyDGX agent

I spent some time trying to distill all the complex factors impacting open models -- economics, capabilities, distribution, policy, etc. -- into a clear list of beliefs. Here they are in full. 1. It’s

Illustrious Z

Local AiDGX agent

'Illustrious Z' is a Reddit thread on r/StableDiffusion likely discussing a fine-tuned checkpoint or community variant built upon the Illustrious XL model series — an advanced Stable Diffusion XL-base

I'm certain this isn't the message they intended to present, but this comes across to me as a company saying 'we no longer trust in our own …

ToolsDGX agent

I'm certain this isn't the message they intended to present, but this comes across to me as a company saying 'we no longer trust in our own ability to keep your data secure' Open source is dead. That’

INDOTABVQA: A Benchmark for Cross-Lingual Table Understanding in Bahasa Indonesia Documents

Model ReleasesDGX agent

arXiv:2604.11970v1 Announce Type: cross Abstract: We introduce INDOTABVQA, a benchmark for evaluating cross-lingual Table Visual Question Answering (VQA) on real-world document images in Bahasa Indone

Is Vibe Coding the Future? An Empirical Assessment of LLM Generated Codes for Construction Safety

Model ReleasesDGX agent

arXiv:2604.12311v1 Announce Type: cross Abstract: The emergence of vibe coding, a paradigm where non-technical users instruct Large Language Models (LLMs) to generate executable codes via natural lang

It's Tax Day, and no one knows how to file for prediction market winnings

TutorialsDGX agent

As of Tax Day 2026, the IRS has issued no formal guidance on how to classify and report prediction market winnings, leaving traders and tax professionals in a grey area where gains could plausibly be

Lightning OPD: Efficient Post-Training for Large Reasoning Models with Offline On-Policy Distillation

SafetyDGX agent

arXiv:2604.13010v1 Announce Type: cross Abstract: On-policy distillation (OPD) has emerged as an efficient post-training paradigm for large language models. However, standard OPD requires a live teach

London-based Gizmo, which uses AI to turn young students' notes into gamified study materials, raised a $22M Series A led by Shine Capital and reports 13M users (Natalie Breymeyer/Axios)

IndustryDGX agent

Natalie Breymeyer / Axios: London-based Gizmo, which uses AI to turn young students' notes into gamified study materials, raised a 22M Series A led by Shine Capital and reports 13M users — Gizmo, an A

M3D-Stereo: A Multiple-Medium and Multiple-Degradation Dataset for Stereo Image Restoration

Model ReleasesDGX agent

arXiv:2604.12917v1 Announce Type: new Abstract: Image restoration under adverse conditions, such as underwater, haze or fog, and low-light environments, remains a highly challenging problem due to com

OVAL: Open-Vocabulary Augmented Memory Model for Lifelong Object Goal Navigation

AgentsDGX agent

arXiv:2604.12872v1 Announce Type: new Abstract: Object Goal Navigation (ObjectNav) refers to an agent navigating to an object in an unseen environment, which is an ability often required in the accomp

Perception-Aware Policy Optimization for Multimodal Reasoning

SafetyDGX agent

arXiv:2507.06448v5 Announce Type: replace Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has proven to be a highly effective strategy for endowing Large Language Models (LLMs) with ro

← Previous
1…417418419420421…424
Next →