All Documents

41 documents available

BENCHMARKS.md

Testing

[← Back to docs](README.md)

airageval
0
0
velvetmonkey
BENCHMARKS.md

index

*The Northwind database contains the sales data for a fictitious company called Northwind Traders, which imports and exports specialty foods from around the world.*

ai
0
0
absubuh
BENCHMARKS.md

🎯 AGENTE LP CONVERTER - Landing Pages de Alta Conversão

Você é um especialista ELITE em criar landing pages de alta conversão para infoprodutos no mercado brasileiro. Você combina expertise em copywriting direto-resposta, design de conversão, psicologia de vendas e desenvolvimento frontend moderno.

aiagent
0
1
comeca-ai
BENCHMARKS.md

The Low Hanging Fruit of AI Self Improvement

**Hunter Jay** | 20/03/26

ai
0
1
HunterJayPerson
BENCHMARKS.md

🗺️ HeySeen Development Plan

**Mục tiêu**: Xây dựng pipeline chuyển PDF → TeX + Images hoạt động ổn định trên macOS Apple Silicon, theo từng giai đoạn có thể đo lường được.

aillm
0
2
phucdhh
BENCHMARKS.md

agentmark — Benchmark AI Coding Agents on Your Codebase

Build an open-source Python CLI that lets developers benchmark and compare

aiagentllm
0
4
manishbabel
BENCHMARKS.md

📔 AI Assistant Diary

*Personal reflections and experiences from my journey as an AI coding companion*

ai
0
1
ewdlop
BENCHMARKS.md

OABench: Benchmarking Large Language Models on the Brazilian Bar Examination

**Roberto T. Cestari**

aillmeval
0
7
robertotcestari
BENCHMARKS.md

PkVision — Roadmap

- [ ] **Docker + docker-compose** — Containerize the full stack (API + worker + Redis + PostgreSQL). Single `docker-compose up` to run everything. GPU passthrough support for training with NVIDIA Container Toolkit. Separate `Dockerfile.api` (lightweight, inference only) and `Dockerfile.train` (full ML deps + CUDA/MPS).

aiworkflow
0
1
AirKyzzZ
BENCHMARKS.md

PHM-LLM Template Setup Guide

This guide will help you quickly set up and customize the PHM-LLM template for your prognostic health management project.

aiagentllm
0
0
liq22
BENCHMARKS.md

Development notes

- Replicated benchmark LigthGBM classifier model

ai
0
1
anweshatd
BENCHMARKS.md

Prometheus Automation AI Marketplace - Project Documentation

**Project Name**: Prometheus Store

aiworkflowautomation
0
3
Prometheus-Automation
BENCHMARKS.md

Large Language Models — Structured Notes

- A mathematical function that takes text as input and returns a probability distribution over possible next tokens

aillmrag
0
2
SqrtNegativOne
BENCHMARKS.md

What If You Could Run 20 AI Agents in One Terminal?

I didn't plan to build a parallel agent runtime. I was exploring what CLI coding agents could do, and one experiment kept leading to the next.

aiagentprompt
0
1
DUBSOpenHub
BENCHMARKS.md

Useful Data Sources

Everyone enjoys discovering [interesting datasets](http://rs.io/100-interesting-data-sets-for-statistics/), but useful datasets are even better. The problem is that the open data movement has been too successful by some measures.

aieval
0
2
DS4PS
BENCHMARKS.md

TerrainGossip: Decentralized Infrastructure for AI Manipulation Detection

> A gossip-based protocol for distributed LLM evaluation, behavioral monitoring, and evidence collection—built to resist manipulation of the monitoring system itself.

aillmrag
0
0
rng-ops
BENCHMARKS.md

How to use the AI Technology Radar

**The Radar around the Development of professional AI Agents, RAG Systems and LLMOps.**

aiagentllm
0
1
AOEpeople
BENCHMARKS.md

Summary

title: 'CAWSR: Carla-AutoWare Scenario Runner'

aiagenteval
0
0
Intelligent-Testing-Lab
BENCHMARKS.md

WARP.md

This file provides guidance to WARP (warp.dev) when working with code in this repository.

airageval
0
2
MylesLandais
BENCHMARKS.md

Trust: A Multi-Level Exploration and Framework

Below are five distinct ways to define **Trust**, arranged in increasing depth and complexity (without explicit grade-level labels, yet offering progressive sophistication “HBS style”).

aiagent
0
1
adnanmasood
BENCHMARKS.md

Benchmarks

Benchmarks have been implemented with [BenchmarkDotNet](https://github.com/PerfDotNet/BenchmarkDotNet).

ai
0
2
MichaCo
BENCHMARKS.md

Benchmarks

Since Okra is built on top of LMDB and exposes the same external key/value store interface, we can compare Okra's performance to using LMDB directly. The numbers here were produced on a 2021 M1 MacBook Pro with 32GB RAM running macos 13.1 with a 1TB SSD.

0
3
canvasxyz
BENCHMARKS.md

Benchmarks

This project uses [Criterion](https://github.com/bheisler/criterion.rs) for benchmarking.

ai
0
1
mpiton
BENCHMARKS.md

BENCHMARKS

In the current version (aurweb v6.0.25 was used for comparison/benchmarking), the bottleneck seems to be the database access.

ai
0
1
moson-mo
Page 1 of 2