Members-Only
Recent Talks & Demos are for members only
You must be an AI Tinkerers active member to view these talks and demos.
Building the AI pipeline for predictionarchive.com
Learn how to build an AI pipeline for prediction parsing, optimizing prompts for accuracy through an evaluation process, demonstrated live from predictionarchive.com.
Predictionarchive.com is an early stage live website that aggregates predictions from across the internet to drive accountability in the public discourse.
I will demo a live “prediction parsing” workflow from the live website and review the architecture that underpins the system.
This website tracks and evaluates predictions on diverse global events.
- LangfuseLangfuse is the open-source LLM engineering platform: gain full observability, manage prompt versions, and run production-grade evaluations.Langfuse delivers the essential LLM engineering stack. It's an open-source platform for full-lifecycle management of your AI applications (agents, chains, etc.). Use its comprehensive tracing (OpenTelemetry-based) to debug complex, non-deterministic interactions and track exact cost/latency metrics. The system provides robust prompt management (versioning, A/B testing) and flexible evaluation tools to measure output quality and monitor production health. Integrations are native: connect with Langchain, OpenAI, and LlamaIndex via Python/JS SDKs for immediate control and clarity over your LLM deployment.
- OpenAI APIOpenAI API: Your direct gateway to cutting-edge AI models (GPT-4o, DALL-E 3, Whisper), enabling scalable, multimodal intelligence integration into any application.The OpenAI API provides authenticated, programmatic access to a powerful suite of generative AI models. Developers leverage REST endpoints and official libraries (Python, Node.js) to integrate capabilities like advanced text generation (GPT-4o), image creation (DALL-E 3), and speech-to-text transcription (Whisper). This platform is engineered for scale, supporting millions of daily requests for tasks from complex reasoning to real-time customer support agents, ensuring your application gets reliable, state-of-the-art intelligence.
- FastAPIFastAPI is a modern, high-performance Python web framework for building APIs with automatic OpenAPI documentation.FastAPI is a robust, high-speed Python web framework: it is built on Starlette (for async capabilities) and Pydantic (for data validation and serialization). Leveraging standard Python 3.8+ type hints, the framework automatically generates interactive API documentation (Swagger UI/ReDoc) and enforces data validation, effectively reducing developer-induced errors by an estimated 40%. This architecture delivers performance on par with Node.js and Go, significantly increasing feature development speed (up to 300% faster). It is production-ready, fully supporting OpenAPI and JSON Schema standards for all API specifications.
- PostgreSQLPostgreSQL (Postgres): The world's most advanced, open-source object-relational database (ORDBMS), built for reliability and extensibility.PostgreSQL is the premier open-source ORDBMS, proven over 35+ years of active development. It adheres strictly to ACID properties (Atomicity, Consistency, Isolation, Durability), ensuring data integrity for mission-critical workloads. Key features include robust SQL compliance, Multi-Version Concurrency Control (MVCC), and superior extensibility (e.g., custom data types, functions in multiple languages). Advanced capabilities like native JSON/JSONB support and the PostGIS extension (geospatial data) make it a powerful, flexible choice for complex enterprise applications.
- k3sK3s is a lightweight, certified Kubernetes distribution that delivers a production-ready cluster in a single 100MB binary.Engineered for Edge and IoT, K3s strips away legacy drivers and alpha features to maintain a footprint under 512MB of RAM. It replaces etcd with SQLite for single-node setups (while supporting external PostgreSQL for HA) and bundles Traefik, Helm, and Flannel into the core process. You get a fully compliant API that boots in seconds, making it the go-to choice for ARM64 clusters and rapid local development.
Related talks
More from the community
From Messy Inputs to Reliable Community Data: Human-in-the-Loop AI
Denver
Learn how to transform messy event and member inputs into structured community data using a human-in-the-loop AI workflow,…
No-code front-ends with AI assisted backend
Seattle
This talk demonstrates building a no-code React frontend with a custom Python/FastAPI AI backend to create animal mashup…
journaling and note-taking with inline AI
San Francisco
Seams Showing: Deconstructionist Approach to AI Engineering
DC
Explore a deconstructionist approach to AI engineering for reliable MTG rules adjudication. This talk details a layered system…
How building with AI showed me where all the design decisions were hiding
Dublin
A non-coder designer built an AI SaaS platform using Claude Code, Codex, and Supabase, detailing agent orchestration, feedback…
Build bulletproof generative AI applications
Berlin
Learn to build fast, reliable generative AI apps. This talk covers caching, fact-checking, and RAG pipelines with real…
Compose Email
Loading recent emails...