AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
All
84,630Total entries
1Added by human
84,629Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,603 results
Model Releases

FinRED: An Expert-Guided Benchmark Generation and Evaluation Framework for Financial LLM Red-Teaming

DGX agent

arXiv:2606.19887v2 Announce Type: replace-cross Abstract: Existing safety benchmarks target general adversarial scenarios but miss finance-specific risks. Financial LLMs face regulatory compliance vio

model-releasesarxiv-cs-ai
25 Jun 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Flexible Gravitational-Wave Parameter Estimation with Transformers

DGX agent

arXiv:2512.02968v2 Announce Type: replace-cross Abstract: Gravitational-wave data analysis relies on accurate and efficient methods to extract physical information from noisy detector signals, yet the

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

FlowID : Enhancing Forensic Identification with Latent Flow-Matching Models

DGX agent

arXiv:2603.29591v2 Announce Type: replace Abstract: Every day, many people die under violent circumstances, whether from crimes, war, migration, or climate disasters. Medico-legal and law enforcement

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

FreeStory: Training-Free Character Consistency for Free-Form Visual Storytelling

DGX agent

arXiv:2606.25079v1 Announce Type: new Abstract: Visual storytelling aims to generate image sequences that are both aligned with narrative prompts and consistent in character appearance across images.

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

From Sounds to Scenes: A Benchmark for Evaluating Context-Aware Auditory Scene Understanding in Large Audio Language Models

DGX agent

arXiv:2606.25391v1 Announce Type: cross Abstract: Recent Large Audio Language Models (LALMs) have achieved remarkable progress in audio perceptual tasks across individual acoustic layers, including sp

model-releasesarxiv-cs-ai
25 Jun 2026
Model Releases

From Sparse and Imperfect 2D Anchors to Consistent 3D Gaussian Street Scenes: Support-Aware Appearance

DGX agent

arXiv:2606.26007v1 Announce Type: new Abstract: Image priors can synthesize target conditions for 3D Gaussian street scenes, but independently edited views do not define a coherent 3D target. Direct f

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

FunPiQ: A New Benchmark for Pixel-Level Quality Assessment in Fundus Images

DGX agent

arXiv:2606.25915v1 Announce Type: new Abstract: Color fundus photography (CFP) is the most common ophthalmic imaging modality for large-scale screening. However, it is highly susceptible to degradatio

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

Gaussian Mean Field Variational Inference can Overestimate Predictive Variance

DGX agent

arXiv:2606.25745v1 Announce Type: cross Abstract: Mean Field Variational Inference (MFVI) is widely understood to underestimate posterior variance. By analysing conjugate Bayesian Linear Regression (B

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

Generating Input Distributions for Explaining Portfolio Optimization Pipelines

DGX agent

arXiv:2606.25808v1 Announce Type: cross Abstract: We propose a predict-optimize-explain framework that uses gradient-based sample generation to interpret various portfolio models by identifying macroe

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

Geometry-Aware Online Scheduling for LLM Serving: From Theoretical Bound to System Practice

DGX agent

arXiv:2606.22327v2 Announce Type: replace Abstract: The explosive demand for interactive Large Language Model serving has highlighted the management of the Key-Value cache's dynamic memory footprint a

model-releasesarxiv-cs-ai
25 Jun 2026
Model Releases

Google launches a Google Finance app for Android, with market data, financial news, and an AI-powered 'Key Moments' feature, and plans an iOS version this year (Aisha Malik/TechCrunch)

DGX agent

Aisha Malik / TechCrunch: Google launches a Google Finance app for Android, with market data, financial news, and an AI-powered “Key Moments” feature, and plans an iOS version this year — Google on Th

model-releasestechmeme
25 Jun 2026
Model Releases

GroundSet: A Cadastral-Grounded Dataset for Spatial Understanding with Vector Data

DGX agent

arXiv:2603.14609v2 Announce Type: replace Abstract: Precise spatial understanding in Earth Observation is essential for translating raw aerial imagery into actionable insights for critical application

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

HG-Bench: A Benchmark for Multi-Page Handwritten Answer-Region Grounding in Automated Homework Assessment

DGX agent

arXiv:2606.25491v1 Announce Type: new Abstract: Automated homework assessment depends not only on recognizing student answers, but also on accurately locating where each answer and each intermediate r

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

Hierarchical Graph Learning for Calendar Spread Strategies in Commodity Futures Markets

DGX agent

arXiv:2606.25811v1 Announce Type: cross Abstract: Commodity futures can be represented hierarchically, with underlying assets at the upper level and individual futures contracts at the lower level. En

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

Hierarchical Reinforcement Learning for Neural Network Compression (HiReLC): Pruning and Quantization

DGX agent

arXiv:2606.26002v1 Announce Type: new Abstract: We present HiReLC, a hierarchical ensemble-reinforcement learning framework for automated joint quantization and structured pruning of deep neural netwo

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

Hitting a Moving Target: Test-Time Adaptation for AI Text Detection under Continual Distribution Shift

DGX agent

arXiv:2606.25152v1 Announce Type: new Abstract: Deployed approaches for AI text detection often rely on training-time access to labeled datasets of both human-written and AI-generated text. This appro

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

How agents are transforming work

DGX agent

OpenAI discusses how AI agents are reshaping workplace processes and productivity by automating complex tasks, handling multi-step workflows, and enabling workers to focus on higher-level decision-mak

model-releasesopenai
25 Jun 2026
Model Releases

How Reliable Is Your Jailbreak Judge? Calibration and Adversarial Robustness of Automated ASR Scoring

DGX agent

arXiv:2606.25487v1 Announce Type: new Abstract: Almost every paper on LLM jailbreaks and prompt injection reports an attack-success rate (ASR), and that number is assigned not by people but by an auto

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

How Robust is OCR-Reasoning? Evaluating OCR-Reasoning Robustness of Vision-Language Models under Visual Perturbations

DGX agent

arXiv:2606.26041v1 Announce Type: cross Abstract: Vision-language models (VLMs) have achieved strong performance on OCR-based benchmarks and increasingly focused on text-rich understanding, but their

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

How Small Can 6G Reason? Scaling Tiny-to-Small Language Models for AI-Native Networks

DGX agent

arXiv:2603.02156v2 Announce Type: replace-cross Abstract: Emerging 6G visions, reflected in ongoing standardization efforts within 3GPP, IETF, ETSI, ITU-T, and the O-RAN Alliance, increasingly charact

model-releasesarxiv-cs-ai
25 Jun 2026
Model Releases

https://openai.com/index/how-agents-are-transforming-work/

DGX agent

OpenAI discusses how AI agents are reshaping workplace productivity and operations, likely covering applications of autonomous AI systems in automating tasks, enhancing decision-making, and transformi

model-releasesopenai--x
25 Jun 2026
Model Releases

I spend 10 minutes a day trying a new app. http://matrix.build is my favorite app today. You can tell from a product whether the team's thin…

DGX agent

I spend 10 minutes a day trying a new app. http://matrix.build is my favorite app today. You can tell from a product whether the team's thinking is clear. The thing I love about matrix is it helps you

model-releaseszhipu-ai--x
25 Jun 2026
Model Releases

I'll be talking more about Claude Tag with @petergyang and at AIE with @_catwu. Let me know if there's anything you'd like us to dive into m…

DGX agent

I'll be talking more about Claude Tag with @petergyang and at AIE with @_catwu. Let me know if there's anything you'd like us to dive into more! Claude Tag is the next evolution of agents. It's a proa

model-releasesthariq--x
25 Jun 2026
Model Releases

Improving Factuality of 3D Brain MRI Report Generation with Paired Image-domain Retrieval and Text-domain Augmentation

DGX agent

arXiv:2411.15490v2 Announce Type: replace Abstract: Acute ischemic stroke (AIS) requires time-critical decision-making, where inaccurate interpretation of neuroimaging findings can lead to irreversibl

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

Improving Zero-Shot Offline RL via Behavioral Task Sampling

DGX agent

arXiv:2604.25496v2 Announce Type: replace Abstract: Offline zero-shot reinforcement learning (RL) aims to learn agents that optimize unseen reward functions without additional environment interaction.

model-releasesarxiv-cs-ai
25 Jun 2026
Model Releases

In a joint Fireworks and @Faros_AI evaluation of 211 real engineering tasks, Claude Code + GLM-5.2 beat both Claude Code + Opus 4.8 and Code…

DGX agent

In a joint Fireworks and @Faros_AI evaluation of 211 real engineering tasks, Claude Code + GLM-5.2 beat both Claude Code + Opus 4.8 and Codex + GPT-5.5: - Judge score: 0.568 vs. 0.521 and 0.466 - Time

model-releasesfireworks-ai--x
25 Jun 2026
Model Releases

In-Context World Modeling for Robotic Control

DGX agent

arXiv:2606.26025v1 Announce Type: cross Abstract: Modern Vision-Language-Action (VLA) models often fail to generalize to novel setups, such as altered camera viewpoints or robot morphologies, because

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

In HF GGUF section of models, we are emphasizing MTP heads with its own sign 𝗠𝗧𝗣

DGX agent

In HF GGUF section of models, we are emphasizing MTP heads with its own sign 𝗠𝗧𝗣 llama.cpp adds MTP for the Qwen3.6 family This is a significant milestone for the local AI ecosystem. The performance j

model-releasesgeorgi-gerganov--x
25 Jun 2026
Model Releases

In this episode, @OpenAI Chief Research Officer @markchen90 joins @allenpark to flambé shrimp, cook Korean stew, and chat about being at the…

DGX agent

In this episode, @OpenAI Chief Research Officer @markchen90 joins @allenpark to flambé shrimp, cook Korean stew, and chat about being at the frontier of AI research: why scaling laws and pre-training

model-releasesswyx--x
25 Jun 2026
Model Releases

IndicContextEval: A Benchmark for Evaluating Context Utilisation in Audio Large Language Models Across 8 Indic Languages

DGX agent

arXiv:2606.19157v2 Announce Type: replace-cross Abstract: AudioLLMs enable speech recognition conditioned on textual prompts such as domain descriptions or entity lists. However, it remains unclear wh

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

Internal Data Repetition Destroys Language Models

DGX agent

arXiv:2606.24998v1 Announce Type: new Abstract: Language models are running out of high-quality training data, and even aggressively deduplicated corpora retain some amount of repetition. Earlier cont

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

Inverse Reinforcement Learning for Interpretable Keystroke Biomarkers in Parkinson's Disease

DGX agent

arXiv:2606.25270v1 Announce Type: new Abstract: Keystroke dynamics have been explored extensively as a passive digital biomarker for Parkinson's disease (PD), typically by extracting summary statistic

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

InvestPhilBench: A Multi-Layer Dynamic Benchmark for Evaluating Large Language Model Procedural Reasoning in Expert Investment Philosophy

DGX agent

arXiv:2606.25984v1 Announce Type: cross Abstract: Large language models are increasingly deployed as investment research assistants, yet no benchmark tests whether they can accurately reconstruct and

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

Invoice Haystack: Benchmarking Document Retrieval and Visual Question Answering Under Strong Visual Homogeneity

DGX agent

arXiv:2606.25343v1 Announce Type: new Abstract: Vision Language Models have achieved near-human performance on single-document Visual Question Answering, yet their effectiveness degrades significantly

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

KidRisk: Benchmark Dataset for Children Dangerous Action Recognition

DGX agent

arXiv:2606.25298v1 Announce Type: new Abstract: Children are naturally energetic, and during their spontaneous activities, they often encounter potentially dangerous situations, especially when lackin

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

Laplace--Fisher Gate Identities for Optimal Matrix-Gated Blended Score Estimation

DGX agent

arXiv:2606.25169v1 Announce Type: cross Abstract: Sampling from an unnormalized target by reversing an Ornstein--Uhlenbeck diffusion requires the score of each noise-perturbed marginal. Tweedie's iden

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

Latent Block-Diffusion Temporal Point Processes: A Semi-Autoregressive Framework for Asynchronous Event Sequence Generation

DGX agent

arXiv:2606.24982v1 Announce Type: new Abstract: Modeling and sampling from the underlying distribution of asynchronous event sequences are crucial in various real-world applications, including social

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

Learning Dynamical Systems from Multiple Sparse Datasets: A Hierarchical Bayesian Modeling Approach

DGX agent

arXiv:2606.24966v1 Announce Type: new Abstract: Estimating parameters of dynamical systems from sparse, noisy, and irregularly sampled data is often severely ill-conditioned. When multiple related dat

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

Learning to Adapt: Reptile-D-Learning for Robust and Efficient Control Under Parametric Uncertainty

DGX agent

arXiv:2606.25659v1 Announce Type: new Abstract: Learning-based Lyapunov Control (LLC) provides formal stability guarantees for nonlinear systems, but its validity relies heavily on accurate system mod

model-releasesarxiv-cs-ro
25 Jun 2026
Model Releases

LEVIRDet: A Million-Scale 159-Category Dataset and Foundation Model for Universal Remote Sensing Object Detection

DGX agent

arXiv:2606.25312v1 Announce Type: new Abstract: Remote sensing object detection has advanced rapidly with the development of large-scale benchmarks and modern detection architectures. However, existin

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

LibEvoBench: Probing Temporal Knowledge Stratification in Code Generation Models

DGX agent

arXiv:2606.25402v1 Announce Type: cross Abstract: Large software projects often depend on older versions of libraries, even as APIs continue to evolve across releases. This creates a challenge for LLM

model-releasesarxiv-cs-ai
25 Jun 2026
Model Releases

LLM Performance on a Real, Double-Marked GCSE Benchmark

DGX agent

arXiv:2606.24973v1 Announce Type: new Abstract: We introduce a dataset of 32,534 double-marked real student responses to GCSE mock exams (GCSEs are the UK's national exams, taken at age ~16), spanning

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

LLMs are getting better at writing GPU kernels. Multi-GPU kernels are the harder test. At @aiDotEngineer World's Fair, @simran_s_arora will …

DGX agent

LLMs are getting better at writing GPU kernels. Multi-GPU kernels are the harder test. At @aiDotEngineer World's Fair, @simran_s_arora will share ParallelKernelBench, an open-source benchmark built fr

model-releasestogether-ai--x
25 Jun 2026
Model Releases

MacroLens: A Multi-Task Benchmark for Contextual Financial Reasoning under Macroeconomic Scenarios

DGX agent

arXiv:2606.24950v1 Announce Type: new Abstract: Financial decision-making is contextual: forecasting prices, valuing companies, and assessing event exposure weigh price history, accounting fundamental

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

Make Watergate Great Again.

DGX agent

Make Watergate Great Again. JD Vance: 'I think Nixon's historical legacy is enjoying a bit of a renaissance, and deservedly so. I joked that if Watergate happened tomorrow, it would be like a 12 hours

model-releasesanthropic--x
25 Jun 2026
Model Releases

MedLayBench-V: A Large-Scale Benchmark for Expert-Lay Semantic Alignment in Medical Vision Language Models

DGX agent

arXiv:2604.05738v2 Announce Type: replace Abstract: Medical Vision-Language Models (Med-VLMs) have achieved expert-level proficiency in interpreting diagnostic imaging. However, current models are pre

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

Memory-Efficient Policy Libraries with Low-Rank Adaptation in Reinforcement Learning

DGX agent

arXiv:2606.25700v1 Announce Type: new Abstract: When fine-tuning Large Language Models (LLMs), there has been success in minimizing both memory usage and computation with Parameter-Efficient Fine-Tuni

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

MINIF2F-DAFNY: LLM-Guided Mathematical Theorem Proving via Auto-Active Verification

DGX agent

arXiv:2512.10187v3 Announce Type: replace Abstract: LLMs excel at reasoning, but validating their steps remains challenging. Formal verification offers a solution through mechanically checkable proofs

model-releasesarxiv-cs-lg
25 Jun 2026
← Previous
1…165166167168169…471
Next →