AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,106 results
Model Releases

FanarGuard: A Culturally-Aware Moderation Filter for Arabic Language Models

DGX agent

arXiv:2511.18852v2 Announce Type: replace Abstract: Content moderation filters are a critical safeguard against alignment failures in language models. Yet most existing filters focus narrowly on gener

model-releasesarxiv-cs-cl
11 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

FEAST: Federated Shared-Space Training for Resource-Heterogeneous Clients

DGX agent

arXiv:2608.09250v1 Announce Type: new Abstract: Federated learning (FL) must serve devices with varying computational capabilities. A fixed model cannot suit all devices, while training one model per

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

FedA2L: Adaptive layer-wise learning rate adjustment in decentralized federated learning

DGX agent

arXiv:2608.09208v1 Announce Type: cross Abstract: Decentralized intelligence systems with heterogeneous devices and limited coordination increasingly rely on decentralized federated learning (DFL). Ho

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

FemWear: A Specialized Wearable Foundation Model for Women's Health

DGX agent

arXiv:2608.08244v1 Announce Type: new Abstract: General wearable foundation models are pretrained across broad sensor streams and populations, but are not designed around women's-health tasks. We intr

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Findings of the First Teaching Monster Challenge: A Benchmark of Pedagogical Content Knowledge in AI Agents

DGX agent

arXiv:2608.08852v1 Announce Type: new Abstract: AI agents can now solve problems, answer like subject experts, and generate long-form multimodal content. However, whether they can adapt a lesson to fi

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Finite Constant Frontiers and Auditable Regret Certificates for Average-Reward Reinforcement Learning

DGX agent

arXiv:2608.07725v1 Announce Type: new Abstract: Average-reward reinforcement-learning regret is known up to logarithmic factors, but the numerical content of published guarantees is difficult to compa

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

FitAQA: A Benchmark of Fitness Action Quality Assessment for Multimodal Large Language Models

DGX agent

arXiv:2608.08736v1 Announce Type: new Abstract: Fitness Action Quality Assessment (AQA) is important for intelligent sports training, yet the capabilities of Multimodal Large Language Models (MLLMs) i

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Floating-Point Neural Network Verification at the Software Level

DGX agent

arXiv:2510.23389v2 Announce Type: replace-cross Abstract: The behaviour of neural network components must be proven correct before deployment in safety-critical systems. Unfortunately, existing neural

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

Flow-by-Flow:Content-Judgment Bypass for Governing AI Output in High-Loss Domains

DGX agent

arXiv:2608.07474v1 Announce Type: new Abstract: Prior work showed that human-in-the-loop oversight becomes structurally untenable in high-loss domains when AI output velocity V exceeds human cognitive

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

FoMoH: A clinically meaningful foundation model evaluation for structured electronic health records

DGX agent

arXiv:2505.16941v4 Announce Type: replace-cross Abstract: Foundation models (FMs) promise to address core limitations of traditional supervised machine learning: (i) reliance on large amounts of label

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

ForestBench: A Unified Graph Framework for Evaluating Multi-Agent Collaboration

DGX agent

arXiv:2608.08605v1 Announce Type: new Abstract: Multi-agent systems (MAS) built on Large Language Models (LLMs) are proliferating rapidly, but their heterogeneous execution traces provide no common ba

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Forgetting-Resistant and Lesion-Aware Source-Free Domain Adaptive Fundus Image Analysis with Vision-Language Model

DGX agent

arXiv:2602.19471v2 Announce Type: replace Abstract: Source-free domain adaptation (SFDA) aims to adapt a model trained in the source domain to perform well in the target domain, with only unlabeled ta

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

FreSH: Frequency-Segmented Hierarchical Multi-Expert Framework for Multivariate Time Series Classification

DGX agent

arXiv:2608.08207v1 Announce Type: cross Abstract: Multivariate Time Series Classification (MTSC) demands models that can effectively capture complex temporal patterns across multiple scales while rema

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

FriskAI launches with $3.6M to show enterprises what their AI agents are doing

DGX agent

Runtime intelligence startup FriskAI Inc. launched today with 3.6 million in pre-seed funding to give enterprises a record of what their artificial intelligence agents actually do once they go into pr

model-releasessiliconangle
11 Aug 2026
Model Releases

From Benchmark Performance to Tool Deployment: Human-in-the-Loop Anomaly Detection

DGX agent

arXiv:2608.07770v1 Announce Type: cross Abstract: Automated anomaly detection methods often report strong performance on curated academic benchmarks, but their behavior under real-world industrial con

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

From Diagnosis to Correction: Benchmarking and Improving Real-World Table Parsing

DGX agent

arXiv:2608.09842v1 Announce Type: new Abstract: Recent document parsers achieve table TEDS scores above 93 on OmniDocBench v1.6, yet community feedback and our audit reveal persistent failures on comp

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

From Evaluated Models to Evaluation Aids: A Multi-Evidence Study of LLM-Based Difficulty Calibration for Programming Examinations

DGX agent

arXiv:2608.07523v1 Announce Type: cross Abstract: Difficulty differences across parallel-class programming examinations affect the fairness of course assessment. This study repositions large language

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

From Manuals to Maintenance: Fine-Tuning MedGemma for Multi-Modal Imaging System Support in Low-Resource Settings

DGX agent

arXiv:2608.08896v1 Announce Type: new Abstract: Imaging device downtime is a major barrier to healthcare delivery in low- and middle-income countries (LMICs), often driven by limited access to special

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

From Product Search to Preference Articulation: The Economics of Agentic Commerce

DGX agent

arXiv:2608.08395v1 Announce Type: cross Abstract: Generative AI is shifting digital commerce from browsing toward agentic search, in which consumers delegate product discovery to AI agents. We compare

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

From Values to Benchmarks: Evaluating Large Language Models for Governmental Use in Dutch

DGX agent

arXiv:2608.09925v1 Announce Type: cross Abstract: Large language models are increasingly being deployed in governmental settings, yet few existing evaluation frameworks jointly reflect the values of p

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Full-bandwidth transformer

DGX agent

arXiv:2608.08888v1 Announce Type: new Abstract: Autoregressive transformers compute along two axes: horizontally across generated tokens, and vertically through model depth. Dense attention gives each

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Fusion Training for Mathematical Generalization in Large Language Models

DGX agent

arXiv:2608.09893v1 Announce Type: cross Abstract: Thinking Mode Fusion (TMF) enables large language models to support both concise responses and long-form reasoning by unifying a non-thinking mode and

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Gaming Without an Attacker: Benchmark Fingerprinting in LLM-Driven Search Under Selection Pressure

DGX agent

arXiv:2608.08722v1 Announce Type: cross Abstract: Benchmarks for systems that are optimized against the evaluation signal measure something different from what they claim. We document this concretely

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

GENCO - A Unified Neural Solver Embedded in a Development Framework for Steady-State Grid Analysis

DGX agent

arXiv:2608.09921v1 Announce Type: new Abstract: Foundation models are transforming business workflows and boosting productivity, yet they remain largely absent from engineering domains such as power s

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

GeoRoute: Geometry-Aware Hybrid Inference for Traffic Future-Frame Prediction

DGX agent

arXiv:2608.09493v1 Announce Type: new Abstract: Long-horizon future-frame prediction is important for autonomous driving, traffic surveillance, and intelligent transportation systems, yet remains chal

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

Goal-oriented Navigation Instruction Generation with Tour Video Priors

DGX agent

arXiv:2608.08596v1 Announce Type: new Abstract: Navigation Instruction Generation (NIG) aims to produce step-by-step natural language instructions for navigation guidance. Existing studies primarily t

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

Goedel-Code-Prover: Hierarchical Proof Search for Open State-of-the-Art Code Verification

DGX agent

arXiv:2603.19329v3 Announce Type: replace-cross Abstract: Large language models (LLMs) can generate plausible code but offer limited guarantees of correctness. Formally verifying that implementations

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Google’s Gemini AI app passes 1 billion monthly active users

DGX agent

Google LLC’s Gemini artificial intelligence app has passed 1 billion monthly active users, making it the 14th product in the company’s history to reach that mark. The company announced the milestone t

model-releasessiliconangle
11 Aug 2026
Model Releases

Governing the KV Cache: Preventing Timing Side-Channel Leakage in Multi-Tenant LLM Inference

DGX agent

arXiv:2608.09225v1 Announce Type: cross Abstract: The key-value (KV) cache is the primary throughput optimization in modern large language model (LLM) inference, enabling prefix reuse across requests.

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Gradient Under Microscope: Benchmarking Resource Utilization of Memory-Efficient Gradient Computation Methods

DGX agent

arXiv:2608.08961v1 Announce Type: new Abstract: AI training's rising resource intensity is straining electricity supplies and carbon budgets, motivating systematic study of memory-efficient training o

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

GraphThink: Graph-Enhanced LLM Thinking for Long-Horizon Embodied Task Planning

DGX agent

arXiv:2608.07905v1 Announce Type: new Abstract: Embodied agents using LLM-based planners often struggle with physical hallucinations, poor generalization to long-horizon tasks, and lack of environment

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

GRASP: Granularity-Aware Region Alignment and Semantic Prototype Learning for Fine-Grained Cross-Modal Understanding in Drone Views

DGX agent

arXiv:2608.09270v1 Announce Type: cross Abstract: Fine-grained cross-modal understanding in drone views is essential for aerial vision-language navigation. However, the inherent wide field of view and

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Harmful Content Is Not Enough: Continuation Framing Moderates In-Context Emergent Misalignment

DGX agent

arXiv:2608.08212v1 Announce Type: new Abstract: In-context learning (ICL) can induce emergent misalignment (EM), where narrow misaligned examples alter answers to unrelated questions. Existing prompts

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

HeatCast: A Benchmark for Neighborhood-Scale LST Forecasting across 124 U.S. Cities

DGX agent

arXiv:2608.07640v1 Announce Type: new Abstract: Land Surface Temperature (LST) is a widely used satellite-derived measure of urban surface heat, but there is no shared benchmark for forecasting it at

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

Hidden Language Consistency Phenomena in Reasoning LLMs

DGX agent

arXiv:2608.08447v1 Announce Type: cross Abstract: Multilingual reasoning models are commonly evaluated by whether they arrive at the correct answer, but not by whether they preserve the intended langu

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Hierarchical Self-Improvement: A Framework for Task-Specific Evolvable Agent Harnesses

DGX agent

arXiv:2608.08466v1 Announce Type: new Abstract: Modern LLM agents are often improved by modifying prompts, tools, or workflows manually, while the executable scaffold surrounding the model---the harne

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

High Fidelity Capture, Reconstruction, and Transfer of Human Demonstrations for Robot-Assisted Bathing

DGX agent

arXiv:2608.09127v1 Announce Type: new Abstract: Despite the demand for robots in high-value clinical tasks like bathing, contemporary systems still lack the safety and reliability required for complex

model-releasesarxiv-cs-ro
11 Aug 2026
Model Releases

High-Layer Attention Pruning with Rescaling

DGX agent

arXiv:2507.01900v3 Announce Type: replace Abstract: Pruning is a highly effective approach for compressing large language models (LLMs), significantly reducing inference latency. However, conventional

model-releasesarxiv-cs-cl
11 Aug 2026
Model Releases

Hit Selection Using SSMD-Based Machine Learning Performance Metrics in High-Throughput Screening Assays

DGX agent

arXiv:2608.07609v1 Announce Type: cross Abstract: High-throughput screening (HTS) assays are central to early-stage drug discovery but are often limited by extreme data sparsity, as primary screens ty

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

HOPPER: Learnable Hop Extraction for Linearized Graph Sequence Models

DGX agent

arXiv:2608.09031v1 Announce Type: new Abstract: Graph neural networks typically propagate information through repeated message-passing layers, coupling the distance over which information travels with

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

How Much Does It Cost to Answer My Question? Benchmarking Cloud VLM-based VQA Systems

DGX agent

arXiv:2608.07861v1 Announce Type: new Abstract: Vision-language models (VLMs) are becoming a practical backend for mobile visual question answering (VQA) systems, enabling smartphones and smart glasse

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

How sensitive do we want AI to be? Socio-communicative competencies of large language models in healthcare

DGX agent

arXiv:2608.07511v1 Announce Type: cross Abstract: Background. Effective clinical practice relies heavily on the socio-communicative skills of medical professionals. Large language models (LLMs) have b

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

How to Ask the AI: A User Perspective Survey for Large Language Model Prompting

DGX agent

arXiv:2608.07494v1 Announce Type: cross Abstract: AI tools like ChatGPT and DeepSeek, powered by Large Language Models (LLMs), allow users to obtain instant and effective content responses simply by t

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

I built a weird, low-power llama.cpp server using an Intel N100 + RTX 5060Ti

DGX agent

Everything started with the sudden death of my old ASRock J1900. While looking for the perfect ITX replacement, I stumbled upon the Chinese CW-NAS-ADLN-K motherboard, which looked perfect on paper: In

model-releasesr-localllama
11 Aug 2026
Model Releases

I gave DeepSeek V4 Flash basic vision by training a 40M connector on 100K examples

DGX agent

I wanted to find out whether a huge text-only MoE could be given basic vision without retraining the language model itself. The short answer is yes. I froze DeepSeek V4 Flash and a 417M-parameter Moon

model-releasesr-localllama
11 Aug 2026
Model Releases

I put Gemma 4 E4B and E2B into an e-reader so I can ask my weird questions and share my thoughts in private directly in app.

DGX agent

Here's how it works in the app: Framework: Runs on LiteRT-LM (like Google's AI Edge). Models: Downloads either the E2B (~2.5 GB) or E4B (~3.6 GB) INT4 quantized models directly from ungated litert-com

model-releasesr-localllama
11 Aug 2026
Model Releases

I ran Muse Glimmer @ 1M context - All tests passed.

DGX agent

Heeeey all! I just completed some fun tests with Muse Glimmer, I thought I'd let you know. In fact, the summary below was written by Muse itself! I ran a 2× DGX Spark cluster and got Meta's day-old Mu

model-releasesr-localllama
11 Aug 2026
Model Releases

I tested the CMP170HX

DGX agent

Lots of rumor and misinfo bouncing around, so I put some of these old mining cards to the test. I used 4 of the 8GB cards, set to 64GB each. Lots of models fit entirely on a single card, and you can a

model-releasesr-localllama
11 Aug 2026
← Previous
1…89101112…461
Next →