AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
25,629 results
10 Apr 2026

Matrix Profile for Anomaly Detection on Multidimensional Time Series

Model ReleasesDGX agent

arXiv:2409.09298v2 Announce Type: replace-cross Abstract: The Matrix Profile (MP), a versatile tool for time series data mining, has been shown effective in time series anomaly detection (TSAD). This

MedRoute: RL-Based Dynamic Specialist Routing in Multi-Agent Medical Diagnosis

AgentsDGX agent

arXiv:2604.06180v1 Announce Type: cross Abstract: Medical diagnosis using Large Multimodal Models (LMMs) has gained increasing attention due to capability of these models in providing precise diagnose

MemReader: From Passive to Active Extraction for Long-Term Agent Memory

SafetyDGX agent

arXiv:2604.07877v1 Announce Type: new Abstract: Long-term memory is fundamental for personalized and autonomous agents, yet populating it remains a bottleneck. Existing systems treat memory extraction

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Mining Electronic Health Records to Investigate Effectiveness of Ensemble Deep Clustering

ApplicationsDGX agent

arXiv:2604.07085v1 Announce Type: new Abstract: In electronic health records (EHRs), clustering patients and distinguishing disease subtypes are key tasks to elucidate pathophysiology and aid clinical

Mitigating Visual Context Degradation in Large Multimodal Models: A Training-Free Decoupled Agentic Framework

Model ReleasesDGX agent

arXiv:2509.23322v2 Announce Type: replace Abstract: With the continuous expansion of Large Language Models (LLMs) and advances in reinforcement learning, LLMs have demonstrated exceptional reasoning c

MM-MoralBench: A MultiModal Moral Evaluation Benchmark for Large Vision-Language Models

Model ReleasesDGX agent

arXiv:2412.20718v2 Announce Type: replace Abstract: The rapid integration of Large Vision-Language Models (LVLMs) into critical domains necessitates comprehensive moral evaluation to ensure their alig

MolmoWeb: Open Visual Web Agent and Open Data for the Open Web

AgentsDGX agent

arXiv:2604.08516v1 Announce Type: new Abstract: Web agents--autonomous systems that navigate and execute tasks on the web on behalf of users--have the potential to transform how people interact with t

Oldest octopus fossil found to not be an octopus

IndustryDGX agent

A 300-million-year-old fossil named *Pohlsepia mazonensis*, originally identified as the world's oldest octopus in 2000, has been reclassified in 2026 as a nautiloid — a relative of the modern naut...

ORACLE-SWE: Quantifying the Contribution of Oracle Information Signals on SWE Agents

AgentsDGX agent

arXiv:2604.07789v1 Announce Type: cross Abstract: Recent advances in language model (LM) agents have significantly improved automated software engineering (SWE). Prior work has proposed various agenti

ParkSense: Where Should a Delivery Driver Park? Leveraging Idle AV Compute and Vision-Language Models

AgentsDGX agent

arXiv:2604.07912v1 Announce Type: new Abstract: Finding parking consumes a disproportionate share of food delivery time, yet no system addresses precise parking-spot selection relative to merchant ent

Personalizing Text-to-Image Generation to Individual Taste

SafetyDGX agent

arXiv:2604.07427v1 Announce Type: new Abstract: Modern text-to-image (T2I) models generate high-fidelity visuals but remain indifferent to individual user preferences. While existing reward models opt

PhyAVBench: A Challenging Audio Physics-Sensitivity Benchmark for Physically Grounded Text-to-Audio-Video Generation

Model ReleasesDGX agent

arXiv:2512.23994v3 Announce Type: replace-cross Abstract: Text-to-audio-video (T2AV) generation is central to applications such as filmmaking and world modeling. However, current models often fail to

PIKA: Expert-Level Synthetic Datasets for Post-Training Alignment from Scratch

Model ReleasesDGX agent

arXiv:2510.06670v2 Announce Type: replace Abstract: High-quality instruction data is critical for LLM alignment, yet existing open-source datasets often lack efficiency, requiring hundreds of thousand

Pistachio: Towards Synthetic, Balanced, and Long-Form Video Anomaly Benchmarks

Model ReleasesDGX agent

arXiv:2511.19474v4 Announce Type: replace-cross Abstract: Automatically detecting abnormal events in videos is crucial for modern autonomous systems, yet existing Video Anomaly Detection (VAD) benchma

Prompt reinforcing for long-term planning of large language models

Model ReleasesDGX agent

arXiv:2510.05921v2 Announce Type: replace Abstract: Large language models (LLMs) have achieved remarkable success in a wide range of natural language processing tasks and can be adapted through prompt

Quantitative Estimation of Target Task Performance from Unsupervised Pretext Task in Semi/Self-Supervised Learning

Model ReleasesDGX agent

arXiv:2508.07299v2 Announce Type: replace-cross Abstract: The effectiveness of unlabeled data in Semi/Self-Supervised Learning (SSL) depends on appropriate assumptions for specific scenarios, thereby

SceneScribe-1M: A Large-Scale Video Dataset with Comprehensive Geometric and Semantic Annotations

Model ReleasesDGX agent

arXiv:2604.07990v1 Announce Type: new Abstract: The convergence of 3D geometric perception and video synthesis has created an unprecedented demand for large-scale video data that is rich in both seman

sciwrite-lint: Verification Infrastructure for the Age of Science Vibe-Writing

Model ReleasesDGX agent

arXiv:2604.08501v1 Announce Type: cross Abstract: Science currently offers two options for quality assurance, both inadequate. Journal gatekeeping claims to verify both integrity and contribution, but

SearchAD: Large-Scale Rare Image Retrieval Dataset for Autonomous Driving

Model ReleasesDGX agent

arXiv:2604.08008v1 Announce Type: new Abstract: Retrieving rare and safety-critical driving scenarios from large-scale datasets is essential for building robust autonomous driving (AD) systems. As dat

See where you can catch us next: https://www.together.ai/events

ToolsDGX agent

Together AI maintains a public events page at [together.ai/events](https://www.together.ai/events) listing its past and upcoming appearances, spanning major industry conferences such as NVIDIA GTC,...

Should we be optimizing for limited compute instead of more parameters? Thoughts?

Local AiDGX agent

"The search did not return the specific Reddit thread. However, I can provide a summary based on what the topic is broadly about within the local-AI/Ollama community context:

SMFD-UNet: Semantic Face Mask Is The Only Thing You Need To Deblur Faces

ApplicationsDGX agent

arXiv:2604.07477v1 Announce Type: new Abstract: For applications including facial identification, forensic analysis, photographic improvement, and medical imaging diagnostics, facial image deblurring

Steering the Verifiability of Multimodal AI Hallucinations

TutorialsDGX agent

arXiv:2604.06714v1 Announce Type: new Abstract: AI applications driven by multimodal large language models (MLLMs) are prone to hallucinations and pose considerable risks to human users. Crucially, su

Tabular GANs for uneven distribution

Model ReleasesDGX agent

arXiv:2010.00638v2 Announce Type: replace-cross Abstract: Generative models for tabular data have evolved rapidly beyond Generative Adversarial Networks (GANs). While GANs pioneered synthetic tabular

TEC: A Collection of Human Trial-and-error Trajectories for Problem Solving

TutorialsDGX agent

arXiv:2604.06734v2 Announce Type: replace Abstract: Trial-and-error is a fundamental strategy for humans to solve complex problems and a necessary capability for Artificial Intelligence (AI) systems o

The Art of Building Verifiers for Computer Use Agents

AgentsDGX agent

arXiv:2604.06240v1 Announce Type: cross Abstract: Verifying the success of computer use agent (CUA) trajectories is a critical challenge: without reliable verification, neither evaluation nor training

The ATOM Report: Measuring the Open Language Model Ecosystem

Model ReleasesDGX agent

arXiv:2604.07190v1 Announce Type: cross Abstract: We present a comprehensive adoption snapshot of the leading open language models and who is building them, focusing on the ~1.5K mainline open models

To dive into building AI agents and your own second brain, join cohort 2 of the AI Agent Mastermind: https://joinaiagentmastermind.com/

AgentsDGX agent

The AI Agent Mastermind is a four-week, live cohort program created by Allie K. Miller — a TIME 100 Most Influential People in AI honoree and the most followed voice in AI business, with nearly 2 ...

Tool-MCoT: Tool Augmented Multimodal Chain-of-Thought for Content Safety Moderation

SafetyDGX agent

arXiv:2604.06205v1 Announce Type: cross Abstract: The growth of online platforms and user content requires strong content moderation systems that can handle complex inputs from various media types. Wh

Toward Memory-Aided World Models: Benchmarking via Spatial Consistency

Model ReleasesDGX agent

arXiv:2505.22976v2 Announce Type: replace-cross Abstract: The ability to simulate the world in a spatially consistent manner is a crucial requirements for effective world models. Such a model enables

Towards Real-world Human Behavior Simulation: Benchmarking Large Language Models on Long-horizon, Cross-scenario, Heterogeneous Behavior Traces

Model ReleasesDGX agent

arXiv:2604.08362v1 Announce Type: new Abstract: The emergence of Large Language Models (LLMs) has illuminated the potential for a general-purpose user simulator. However, existing benchmarks remain co

Transforming the Voice of the Customer: Large Language Models for Identifying Customer Needs

TutorialsDGX agent

arXiv:2503.01870v2 Announce Type: replace Abstract: Identifying customer needs (CNs) is fundamental to product innovation and marketing strategy. Yet for over thirty years, Voice-of-the-Customer (VOC)

Vision-Language Foundation Models for Comprehensive Automated Pavement Condition Assessment

ApplicationsDGX agent

arXiv:2604.08212v1 Announce Type: new Abstract: General-purpose vision-language models demonstrate strong performance in everyday domains but struggle with specialized technical fields requiring preci

Vision-Language Navigation for Aerial Robots: Towards the Era of Large Language Models

Model ReleasesDGX agent

arXiv:2604.07705v1 Announce Type: new Abstract: Aerial vision-and-language navigation (Aerial VLN) aims to enable unmanned aerial vehicles (UAVs) to interpret natural language instructions and autonom

Watching @awnihannun at @ollama

Local AiDGX agent

I was unable to retrieve the specific content from the X (Twitter) URL provided (`https://x.com/twid/status/2042425382859841926`), as it is a social media post that requires authentication to acces...

We had Lin on stage: 'the future is millions of models — one per application, one per use case.' Jet delivered a masterclass on reinforcemen…

ToolsDGX agent

We had Lin on stage: 'the future is millions of models — one per application, one per use case.' Jet delivered a masterclass on reinforcement fine-tuning. Rob joined @WorkOS for some hot takes on the

When Personalization Tricks Detectors: The Feature-Inversion Trap in Machine-Generated Text Detection

Model ReleasesDGX agent

arXiv:2510.12476v2 Announce Type: replace Abstract: Large language models (LLMs) have grown more powerful in language generation, producing fluent text and even imitating personal style. Yet, this abi

9 Apr 2026

Managed Agents

ApplicationsHuman

Anthropic's new hosted service for long-running AI agents, designed to solve the challenge of creating systems that support 'programs as yet unthought of.' It abstracts infrastructure management to en

Another banger paper from Microsoft. Why it's a big deal: It teaches reasoning models to compress their own chain-of-thought mid-generation.…

AgentsDGX agent

Another banger paper from Microsoft. Why it's a big deal: It teaches reasoning models to compress their own chain-of-thought mid-generation. The most interesting finding isn't the 2-3x memory savings

Anthropic built a model too risky to release

IndustryDGX agent

Anthropic's new frontier model, Claude Mythos, is the first model the company has publicly deemed too high-risk for general release , due to its advanced cybersecurity capabilities — including the ...

Appknox launches KnoxIQ to prioritize real-world exploitability in AI-driven application security

Model ReleasesDGX agent

Mobile app security solutions provider Appknox today announced the launch of KnoxIQ, an artificial intelligence-native vulnerability assessment capability that introduces a new prioritization and reme

dude is on some generational run. highly recommend reading this anyone into harness design and sourcing evals. and viv is genius in making s…

IndustryDGX agent

The referenced tweet (from @himanshustwts) is not publicly accessible without a login, so its exact content cannot be retrieved. However, based on contextual search results, this post appears to be...

Here is my free workshop on building multi-agent systems, I presented at the @aiDotEngineer London conference together with @Whats_AI. It ha…

Model ReleasesDGX agent

Here is my free workshop on building multi-agent systems, I presented at the @aiDotEngineer London conference together with @Whats_AI. It has code, slides and soon a 2-hour video diving deep into how

I just gave a workshop at @aiDotEngineer in London on building real multi-agent systems. The best part was hearing people laugh, interrupt u…

Model ReleasesDGX agent

I just gave a workshop at @aiDotEngineer in London on building real multi-agent systems. The best part was hearing people laugh, interrupt us with questions.. You could feel they were following, think

If you want an example of what this looks like in practice, check out the '/research-docs' skill I created for Claude Code https://x.com/jer…

Model ReleasesDGX agent

If you want an example of what this looks like in practice, check out the '/research-docs' skill I created for Claude Code https://x.com/jerryjliu0/status/2041564207750246904?s=20 I built a Claude Cod

Is ChatGPT really as bad for the environment as people say?

IndustryDGX agent

The environmental impact of individual ChatGPT use is widely considered overstated: Epoch AI estimated a typical ChatGPT query uses just 0.3 Wh of electricity — ten times less than older 2023 esti...

Judging by my tl there is a growing gap in understanding of AI capability. The first issue I think is around recency and tier of use. I thin…

Model ReleasesDGX agent

Judging by my tl there is a growing gap in understanding of AI capability. The first issue I think is around recency and tier of use. I think a lot of people tried the free tier of ChatGPT somewhere l

Meta's new AI can predict your brain better than a brain scan. TRIBE v2 is a foundation model trained on 1,000+ hours of brain imaging data …

IndustryDGX agent

Meta's new AI can predict your brain better than a brain scan. TRIBE v2 is a foundation model trained on 1,000+ hours of brain imaging data from 720 people. You feed it a video, sound clip, or text, a

My considered opinion is that The Mythos stuff was mostly a myth. Take it as a serious warning sign that we need to get our act together wit…

SafetyDGX agent

My considered opinion is that The Mythos stuff was mostly a myth. Take it as a serious warning sign that we need to get our act together with respect to cybersecurity. But don’t take the details serio

RFK Jr. rewrites CDC panel's charter, opening door to anti-vaccine quacks

IndustryDGX agent

Following a courtroom defeat, HHS Secretary RFK Jr. rewrote the charter of the CDC's Advisory Committee on Immunization Practices (ACIP), broadening its membership criteria, increasing its focus on...

// Scaling Coding Agents via Atomic Skills // Most coding agents train end-to-end on full tasks like resolving GitHub issues. But complex so…

AgentsDGX agent

// Scaling Coding Agents via Atomic Skills // Most coding agents train end-to-end on full tasks like resolving GitHub issues. But complex software engineering is really a composition of simpler skills

Studying Sutton and Barto's RL book and its connections to RL for LLMs (e.g., tool use, math reasoning, agents, and so on)? [D]

AgentsDGX agent

A Reddit discussion thread on r/MachineLearning in which practitioners explore how foundational concepts from Sutton and Barto's *Reinforcement Learning: An Introduction* — including MDPs, policy g...

What’s new in Microsoft Foundry | March 2026

Model ReleasesDGX agent

March ships Foundry Agent Service GA with private networking, GPT-5.4 and GPT-5.4 Mini, Priority Processing, Phi-4 Reasoning Vision, SDK 2.0 GA across Python, JS/TS, Java, and .NET, Fireworks AI and

8 Apr 2026

[AINews] Anthropic @ $30B ARR, Project GlassWing and Claude Mythos Preview — first model too dangerous to release since GPT-2

Model ReleasesDGX agent

Anthropic announced it has grown its ARR from $19B to $30B in just one month, coinciding with the formal unveiling of Claude Mythos Preview — described in leaked company documents as 'by far the m...

Another banger article from the @LangChain team! Harness evolution combined with specialist local models will be the way forward undoubtedly…

AgentsDGX agent

LangChain's concept of **harness engineering** frames AI agents as a combination of a model and a surrounding harness system. An agent equals a model plus a harness — harness engineering is how sy...

Director James Cameron on why Big Tech owning AGI is scarier than any science fiction he's ever made: 'AGI will not emerge from a government…

SafetyDGX agent

Director James Cameron on why Big Tech owning AGI is scarier than any science fiction he's ever made: 'AGI will not emerge from a government funded program. It will emerge from one of the tech giants

Introducing the Child Safety Blueprint

SafetyDGX agent

OpenAI's Child Safety Blueprint, released in April 2026, is a policy framework aimed at combating the rise of AI-enabled child sexual exploitation by combining legal, operational, and technical app...

Meta's new model is Muse Spark, and meta.ai chat has some interesting tools

Model ReleasesDGX agent

Meta announced Muse Spark today, their first model release since Llama 4 almost exactly a year ago. It's hosted, not open weights, and the API is currently 'a private API preview to select users', but

Self-improving agents isn’t a single algorithm - it’s a systems engineering problem involving: - eval data curation + maintenance - experime…

AgentsDGX agent

Self-improving agents isn’t a single algorithm - it’s a systems engineering problem involving: - eval data curation + maintenance - experiment design to battle overfitting - an update algorithm - huma

Some sober thinking about Mythos (full version with links at my newsletter): 1It’s probably not as bad as they say, as AI and cybersecurity …

Model ReleasesDGX agent

Some sober thinking about Mythos (full version with links at my newsletter): 1It’s probably not as bad as they say, as AI and cybersecurity expert @HeidyKhlaaf explains elsewhere (in a thread “As some

← Previous
1…425426427428
Next →