AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
25,639 results
27 Jul 2026

創業以来、オープンソースコミュニティから多くを学び、また研究成果の公開を通じてそこに貢献してきました。オープンなエコシステムが健全なAI産業と技術主権を支える重要な基盤の一つであると考えており、その発展を支持します。 このたび、Sakana AIは、オープンウェイトAIモデルに関…

SafetyDGX agent

創業以来、オープンソースコミュニティから多くを学び、また研究成果の公開を通じてそこに貢献してきました。オープンなエコシステムが健全なAI産業と技術主権を支える重要な基盤の一つであると考えており、その発展を支持します。 このたび、Sakana AIは、オープンウェイトAIモデルに関する公開書簡 「Open Weights and American AI Leadership」に署名しました。 書簡は

An opinionated guide to which AI to use to do stuff

Model ReleasesDGX agent

An opinionated guide to which AI to use to do stuff It's interesting watching the evolution of Ethan Mollick's guide over time. A year ago it was still all about chat - ChatGPT, Claude, Gemini - with

Attackers have frontier AI. Defenders need a frontier AI ecosystem—the best open and closed models, force-multiplied by a global community. …

HardwareDGX agent

Attackers have frontier AI. Defenders need a frontier AI ecosystem—the best open and closed models, force-multiplied by a global community. During the Hugging Face incident, closed AI blocked essentia

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Big update: Among open-weight models, Kimi K3 (Max) is #1 in the Agent Arena with +9.75% net-improvement, surpassing GLM-5.2 (Max) at +7.12%…

Model ReleasesDGX agent

Big update: Among open-weight models, Kimi K3 (Max) is #1 in the Agent Arena with +9.75% net-improvement, surpassing GLM-5.2 (Max) at +7.12%, and landed the #1 spot across 5 signals (see below). Kimi

Constraint-Driven Synthesis of Hyper Petri Nets

SafetyDGX agent

arXiv:2607.22062v1 Announce Type: cross Abstract: This paper addresses the modeling and synthesis of constrained robotic system behaviors using Petri nets (PNs). It investigates how to construct model

Correlating Cross-Iteration Noise for DP-SGD using Model Curvature

TutorialsDGX agent

arXiv:2510.05416v3 Announce Type: replace Abstract: Differentially private stochastic gradient descent (DP-SGD) offers the promise of training deep learning models while mitigating many privacy risks.

Enough is as good as a feast: A Comprehensive Analysis of How Reinforcement Learning Mitigates Task Conflicts in LLMs

Model ReleasesDGX agent

arXiv:2607.22039v1 Announce Type: new Abstract: Model merging plays a crucial role in consolidating multiple specialized models into a single, unified model, especially in the era of large language mo

From Seasonality to Semantics: Benchmarking a Hybrid Probabilistic Forecasting System for Roadblocks in Bolivia

Local AiDGX agent

arXiv:2607.21785v1 Announce Type: cross Abstract: Roadblocks in Bolivia are a social conflict phenomenon with devastating economic impacts, estimated at losses equivalent to 4% of the national Gross D

GLI-AL: A Multi-Modal Glioma MRI Label Resource with Unified Anatomy-Lesion Labels

Model ReleasesDGX agent

arXiv:2607.22135v1 Announce Type: new Abstract: Existing BraTS-GLI datasets provide a widely used benchmark for adult glioma MRI segmentation, but their task definition focuses on tumor subregions and

Great technical paper from Google. Great read on why context beats scale for agents working against unfamiliar APIs. (bookmark it) GPU kerne…

Model ReleasesDGX agent

Great technical paper from Google. Great read on why context beats scale for agents working against unfamiliar APIs. (bookmark it) GPU kernel optimization has KernelBench to hillclimb on. TPUs had not

Great technical paper from Harvard and MIT. It's on role drift in compound LLM systems. (bookmark it) End-to-end RL improves the accuracy of…

SafetyDGX agent

Great technical paper from Harvard and MIT. It's on role drift in compound LLM systems. (bookmark it) End-to-end RL improves the accuracy of a multi-module LLM pipeline without constraining how the mo

I ran the 35B agentic comparison someone asked for (stock vs Ornith vs KAT-Coder, 120 runs)

Model ReleasesDGX agent

Someone in the comments of my 27B post-train bakeoff asked for the 35B version, so I ran it. Same setup as last time: fresh Coder workspaces on my k8s cluster, each driving my own agent (Hermes) headl

If you look at the current lead times for ~$5bn per year of cutting edge GPUs you can probably figure out the time plus one training run to …

HardwareDGX agent

If you look at the current lead times for ~$5bn per year of cutting edge GPUs you can probably figure out the time plus one training run to IlyAGI We are announcing a long-term strategic partnership w

Improving Large Vision-Language Models' Understanding for Flow Field Data

Model ReleasesDGX agent

arXiv:2507.18311v3 Announce Type: replace Abstract: Large Vision-Language Models (LVLMs) have shown impressive capabilities across a range of tasks that integrate visual and textual understanding, suc

In partnership with @Kimi_Moonshot, we now have K3 on Together APIs with capacity for hundreds of millions of TPM on launch day!

AgentsDGX agent

In partnership with @Kimi_Moonshot, we now have K3 on Together APIs with capacity for hundreds of millions of TPM on launch day! Kimi K3 is now live on Together AI. We’re proud to be a Day 0 launch pa

Kimi K3 is now live on @togethercompute! Happy to have Together AI as our day0 launch partner, giving developers immediate access to K3 thro…

AgentsDGX agent

Kimi K3 is now live on @togethercompute! Happy to have Together AI as our day0 launch partner, giving developers immediate access to K3 through high-throughput inference optimized for coding agents an

@Kimi_Moonshot K3 on Together AI is built for long-running agent workflows: → 2.8T parameters and a 1M context window → Native vision for sc…

Model ReleasesDGX agent

@Kimi_Moonshot K3 on Together AI is built for long-running agent workflows: → 2.8T parameters and a 1M context window → Native vision for screenshot-guided coding → Repository navigation and terminal

Looking forward to chatting with @hugobowne today (July 27) at 4 pm PT on the Vanishing Gradient livestream on YouTube. Will cover open sour…

AgentsDGX agent

Looking forward to chatting with @hugobowne today (July 27) at 4 pm PT on the Vanishing Gradient livestream on YouTube. Will cover open source, the newest LLMs & trends, agent frameworks, and whatever

Nice little insights on doing autoresearch with coding agents. Hand a coding agent a dataset, an eval script, one editable file, and no supe…

Model ReleasesDGX agent

Nice little insights on doing autoresearch with coding agents. Hand a coding agent a dataset, an eval script, one editable file, and no supervision. That's autoresearch and it tries to optimize the nu

Opaque Epistemic Mediation: How LLM Deployment Configurations Shape the Validation of Pseudo-Science

Model ReleasesDGX agent

arXiv:2607.22513v1 Announce Type: cross Abstract: Commercial large language models are increasingly used as knowledge references, yet their stance on contested scientific claims is neither stable nor

Proud to join @nvidia as a founding member of the Open Secure AI Alliance. Defenders need open tools, collaboration, and transparency to sec…

HardwareDGX agent

Proud to join @nvidia as a founding member of the Open Secure AI Alliance. Defenders need open tools, collaboration, and transparency to secure AI agents and systems. And customers want the ability to

Six Agent Harness Capabilities for Higher Model Performance

HardwareDGX agent

The performance of AI agents depends not only on the underlying models but also heavily on their “harness”—the surrounding architecture that supplies context, state management, action execution, and t

Teachy Mini: Development and Preliminary Evaluation of a Knowledge-Based Generative Social Robot for Higher Education

ApplicationsDGX agent

arXiv:2607.22345v1 Announce Type: new Abstract: Generative social robots (GSRs) powered by large language models offer new possibilities for personalized tutoring in higher education, but also introdu

Want to go deeper? Join Moonshot AI and Together AI for a technical webinar on how K3 was built and how to use it for production agent workf…

AgentsDGX agent

Together AI has released the Kimi K3 model on its platform as a Day‑0 launch partner for Moonshot AI’s open frontier agentic model, which supports long‑running workflows across code, tools, vision and

26 Jul 2026

90 agentic bakeoff runs: ThinkingCap vs Fable Fusion vs stock Qwen3.6-27B

Model ReleasesDGX agent

Last week someone here said ThinkingCap and Fable Fusion 'really do beat the OG' for agentic work, so I ran it: 6 self-grading tasks, 5 reps, 3 models, 90 isolated runs. Tooling, since that's half the

I built an open-source Ollama canvas where the wires are the actual context

Local AiDGX agent

Most graph-based LLM interfaces use a canvas as a visual layer over what is still a linear chat. I wanted the graph itself to determine what Ollama receives. ThoughtDAG has one rule: wires are the con

“The first autonomous agent cyberattack is an unprecedented event. It deserves an unprecedented response!”

AgentsDGX agent

“The first autonomous agent cyberattack is an unprecedented event. It deserves an unprecedented response!” In the spirit of transparency, here’s what I asked @OpenAI: • Radical transparency: let’s rel

25 Jul 2026

Local alternative to Kling AI 3.0 Motion Control (ComfyUI, 16GB VRAM)

Local AiDGX agent

Hi everyone, I'm looking for a local alternative to Kling AI 3.0 Motion Control that I can run in ComfyUI. What I'm specifically looking for is a model or workflow that allows me to: - Control charact

Mobile Offline LLMs: What do you use them for?

Model ReleasesDGX agent

I've spent the last year or so playing around with open source MLX and GGUF models on iPhone hardware. Given the limitations in memory, GPU/CPU/ANE, and in turn the context window I've been trying to

24 Jul 2026

A Unified Moral-Value Dataset for Instruction Tuning

SafetyDGX agent

arXiv:2607.21279v1 Announce Type: new Abstract: Large language models (LLMs) have developed rapidly and become valuable tools in everyday life. However, how to align LLMs to a particular set of human

Achieving Text-based Person Retrieval with Any Granularity

Model ReleasesDGX agent

arXiv:2607.21057v1 Announce Type: new Abstract: Text-based person retrieval faces a critical but under-explored challenge: the inherent uncertainty of query granularity in real-world scenarios. This p

ADABORD: a novel AdaBoost approach for ordinal classification

Model ReleasesDGX agent

arXiv:2607.21003v1 Announce Type: new Abstract: Ordinal Classification (OC) deals with classification tasks where the classes follow a natural order. Despite the progress in OC, many existing approach

// Agentic Context Management // Great read for the weekend. (bookmark it) Production agents fail less on reasoning and more on what sits in…

AgentsDGX agent

// Agentic Context Management // Great read for the weekend. (bookmark it) Production agents fail less on reasoning and more on what sits in their context. Conversation history, big prompts, huge tool

AI Engineer Paris is back! After an incredible first edition, we’re excited to announce AI Engineer Paris 2026, this time hosted by our frie…

ToolsDGX agent

AI Engineer Paris is back! After an incredible first edition, we’re excited to announce AI Engineer Paris 2026, this time hosted by our friends at @MistralAI. 🎤 CFP is open. Submit your talk by July 3

AppWorld-UL: Benchmarking Diverse Agent-User Interactions for Tool-Use

Model ReleasesDGX agent

arXiv:2607.20536v1 Announce Type: new Abstract: Tool-use agents that address day-to-day digital tasks such as ordering groceries must not only operate applications, but also interact with the user, e.

Backpropagation-Free Test-Time Adaptation for Lightweight EEG-Based Brain-Computer Interfaces

ApplicationsDGX agent

arXiv:2601.07556v2 Announce Type: replace-cross Abstract: Electroencephalogram (EEG)-based brain-computer interfaces (BCIs) face significant deployment challenges due to inter-subject variability, sig

Benchmarking the Personalization Capabilities of Large Language Models

Model ReleasesDGX agent

arXiv:2607.20471v1 Announce Type: new Abstract: Personalization, the act of varying a message to induce action from a specific receiver while keeping sender, channel, and time fixed, has a long tradit

Break Through the Compression Bottleneck: From Theory to Practice

Model ReleasesDGX agent

arXiv:2607.20434v1 Announce Type: cross Abstract: As the parameter size of language models continues to grow, effective model compression is required to reduce their computational and memory overhead.

Can LLMs solve mazes?

Model ReleasesDGX agent

https://reddit.com/link/1v5rvuq/video/bgmwc754i9fh1/player My goal was to create a benchmark to measure the spatial awareness and memory of models. Eventually, I came up with the simple idea of a maze

EmoSpace: Immersive Affective Image Generation Guided by Fine-Grained Emotion Prototypes

SafetyDGX agent

arXiv:2602.11658v2 Announce Type: replace Abstract: Immersive affective content generation aims to create visually compelling VR imagery with controllable emotional nuance, yet existing methods typica

Evaluating the Effectiveness of Persona Simulation in Opinion Prediction with GPT-4.1

Model ReleasesDGX agent

arXiv:2607.20589v1 Announce Type: new Abstract: Persona simulation involves utilizing large language models (LLMs) to anticipate human choices or interactions based on specific characteristic informat

Exploiting Exogenous Structure for Sample-Efficient Reinforcement Learning

AgentsDGX agent

arXiv:2409.14557v4 Announce Type: replace-cross Abstract: We study a structured class of Markov Decision Processes, known as Exo-MDPs, in which the state space is partitioned into exogenous and endoge

From Static Bibliometrics to Dynamic Knowledge Graphs: An LLM-Powered Framework for Modernizing Science, Technology, and Innovation (STI) Analytics

SafetyDGX agent

arXiv:2607.21327v1 Announce Type: cross Abstract: Bibliometric indicators - citation counts, h-indexes, co-authorship networks - have long anchored science, technology, and innovation (STI) analytics,

Generative AI and Agency in Education: A Critical Scoping Review and Thematic Analysis

SafetyDGX agent

arXiv:2411.00631v2 Announce Type: replace-cross Abstract: This scoping review examines the relationship between Generative AI (GenAI) and agency in education, analyzing the literature available throug

Generative Artificial Intelligence in Bioinformatics: A Systematic Review of Models, Applications, and Methodological Advances

SafetyDGX agent

arXiv:2511.03354v2 Announce Type: replace-cross Abstract: Generative artificial intelligence (GenAI) is transforming bioinformatics by advancing genomics, proteomics, transcriptomics, structural biolo

GigaPath-Flash and GigaTIME-Flash: Efficient Pathology Foundation Models for Whole-Slide and Tumor Microenvironment Analysis

Model ReleasesDGX agent

arXiv:2607.18218v2 Announce Type: replace-cross Abstract: Foundation models have emerged as a driving force in computational pathology, with the potential to transform cancer diagnosis, prognosis, and

GroupVideo: Multi-Identity Customized Text-to-Video Generation

SafetyDGX agent

arXiv:2607.21027v1 Announce Type: new Abstract: Current identity customized video generation methodologies are predominantly limited to single-identity scenarios, as the lack of explicit identity sepa

HyperImageNet: A Large-Scale High-Spatial Resolution Hyperspectral Imagery Classification Benchmark

Model ReleasesDGX agent

arXiv:2607.21050v1 Announce Type: new Abstract: We present HyperImageNet, a large-scale benchmark for fine-grained hyperspectral land-cover understanding. The dataset contains 26,084 airborne hyperspe

i2Nav-Robot: A Large-Scale Indoor-Outdoor Robot Dataset for Multi-Sensor Fusion Navigation

AgentsDGX agent

arXiv:2508.11485v3 Announce Type: replace Abstract: Accurate and reliable navigation is crucial for autonomous unmanned ground vehicles (UGVs). However, current UGV datasets fall short in meeting the

ImplicitBBQ: Benchmarking Implicit Bias in Large Language Models through Characteristic Based Cues

Model ReleasesDGX agent

arXiv:2604.01925v2 Announce Type: replace-cross Abstract: Large Language Models increasingly suppress biased outputs when demographic identity is stated explicitly, yet may still exhibit implicit bias

Learn2Zinc: Fine-tuning Small Language Models for Text-to-Model Translation in MiniZinc

Model ReleasesDGX agent

arXiv:2607.20456v1 Announce Type: cross Abstract: Large language models excel at code generation for mainstream programming languages but struggle with rare, domain-specific languages such as MiniZinc

Lessons and Open Questions from a Unified Study of Camera-Trap Species Recognition Over Time

Model ReleasesDGX agent

arXiv:2603.20509v2 Announce Type: replace Abstract: Camera traps are vital for large-scale biodiversity monitoring, yet accurate automated analysis remains challenging due to diverse deployment enviro

Leveraging Biokinetic Knowledge Priors for Data-Scarce Bioprocess Modeling

TutorialsDGX agent

arXiv:2607.20539v1 Announce Type: cross Abstract: While deep learning has accelerated drug discovery, its impact on biomanufacturing has been considerably more limited. The reason is data scarcity. Bi

M^3-Gen: Interpretable Multimodal Generation of Gene Expression Profiles Using Clinical and Imaging Data

TutorialsDGX agent

arXiv:2607.21343v1 Announce Type: cross Abstract: Integrating heterogeneous biomedical data, including clinical metadata, histopathology images, and molecular profiles, is crucial for comprehensive di

Meta is making its AI chatbot more like an assistant

Model ReleasesDGX agent

Meta is upgrading its AI chatbot with new productivity features in a bid to compete with rivals like Gemini, ChatGPT, and Claude. The update will allow Meta AI to tap into your calendar to help you pl

Open Source Tax Engine outperforming fable 5 and gpt sol

Model ReleasesDGX agent

This is an open source and free tax engine which scored 96% on TaxCalcBench [highest ever recorded score till date] surpassing fable 5 and sol with just sonnet 5. The only 2 cases where it missed, it

OpenForgeRL: Train Harness-native Agents in Any Environment

Model ReleasesDGX agent

arXiv:2607.21557v1 Announce Type: new Abstract: Modern AI agents rely on elaborate inference harnesses such as Claude Code, Codex, and OpenClaw to drive multi-turn reasoning, tool use, and access to e

Rushes: A Human Preference Dataset for Pluralistic Alignment

Model ReleasesDGX agent

arXiv:2607.20767v1 Announce Type: new Abstract: We introduce Rushes, a dataset and benchmark for studying revealed human engagement preferences in interactive narrative environments. Rushes is collect

Skill-Contracted Agents for Evidence-Aware Materials Literature Analysis

AgentsDGX agent

arXiv:2607.20431v1 Announce Type: new Abstract: Materials science literature analysis requires simultaneous attention to composition, processing, characterization, and property relationships, yet conv

SOAP, Muon, and Beyond: Pushing LLM Pretraining Scales

Model ReleasesDGX agent

arXiv:2607.20548v1 Announce Type: cross Abstract: Higher-order optimizers such as Muon and SOAP offer faster convergence than AdamW, but their computational cost and numerical stability challenges hav

← Previous
1…378379380381382…428
Next →