Members-Only
Recent Talks & Demos are for members only
You must be an AI Tinkerers active member to view these talks and demos.
Build Your Book: Engineering Sci-Fi with an AI Software Mindset
Learn to build a sci-fi novel using Spec-Driven Development and AI, creating an interactive experience with an ARG and unlocking digital assets from the text.
Traditional writing is linear; software is modular. I applied Spec-Driven Development to architect a Sci-Fi techno-thriller, using AI as a precision execution layer for my lore and scenarios.
I’ll break down the technical workflow used to publish the novel and the custom application built alongside it. We’ll explore how to gamify the reader’s journey through an Alternate Reality Game (ARG), where clues hidden in the text act as keys to unlock digital assets, effectively turning a standalone novel into an interactive platform.
Initializes a stable, low-latency node for 2036 archive recovery operations.
This project applies spec-driven software engineering to AI-assisted narrative architecture.
- Gemini ChatGoogle’s multimodal AI assistant built on the Gemini 1.5 Pro and Flash models for reasoning, coding, and creative collaboration.Gemini Chat serves as the primary interface for Google’s most capable AI models, including Gemini 1.5 Pro with its massive 1-million-token context window. It handles complex workflows across text, code, images, and audio, allowing users to analyze 1,500-page documents or debug Python scripts in seconds. By integrating directly with Google Workspace (Docs, Gmail, Drive), it converts prompts into actionable outputs like project briefs or email drafts. The platform operates on a global scale, supporting over 40 languages and providing real-time information retrieval through Google Search.
- Gemini CLIGemini CLI is your open-source AI agent: it brings the power of Gemini models directly into your terminal.The Gemini CLI is an open-source, terminal-first AI agent providing direct, lightweight access to the Gemini model capabilities. It operates as an interactive Read-Eval-Print Loop (REPL), enabling developers to execute complex tasks like code generation, project analysis, and bug fixing through natural language prompts. The CLI leverages a Reason and Act (ReAct) loop, orchestrating built-in tools (file system operations, shell commands, web search) and custom extensions via the Model Context Protocol (MCP). Users can access powerful models, including Gemini 2.5 Pro and Gemini 3 Pro, for high-performance agentic coding and streamlined workflows, all managed from the command line.
- OpenAI Image-2OpenAI's natively multimodal generator that integrates agentic reasoning to produce pixel-perfect typography and complex 4K compositions.OpenAI Image-2 (internally gpt-image-2) represents a fundamental shift from traditional diffusion models to a natively multimodal architecture integrated directly into the GPT-4o ecosystem. Launched in April 2026, the model introduces a dedicated thinking mode that allows it to plan and reason through spatial relationships before rendering, resulting in a 95% accuracy rate for complex multilingual text. It supports native 2K resolution with 4K upscaling and handles production-ready tasks like consistent character rendering and intricate UI mockups. By moving away from standalone plugins, Image-2 achieves superior instruction following and provides developers with a high-fidelity API capable of generating up to eight coherent outputs from a single prompt.
- Nano BananaNano Banana (Gemini 2.5 Flash Image): The AI image model delivering unprecedented character consistency and multi-image fusion with precise natural language commands.Nano Banana is the viral codename for Google's Gemini 2.5 Flash Image model: a next-generation AI editor. This technology excels at complex, context-aware image manipulation, moving beyond simple generation. Key performance metrics include achieving 95% character consistency across transformations and seamless multi-image fusion (up to 3 separate images). Access the model via the Gemini app or Google AI Studio to execute professional-grade edits, like transforming portraits or generating photorealistic product mockups, all through a single, precise natural language prompt.
- ChatGPTOpenAI's Generative Pre-trained Transformer (GPT) model: a conversational AI chatbot for instant text generation, coding assistance, and complex problem-solving.Launched by OpenAI in November 2022, ChatGPT is a state-of-the-art conversational AI, built on the Generative Pre-trained Transformer (GPT) architecture (e.g., GPT-4). The system is fine-tuned using Reinforcement Learning from Human Feedback (RLHF) to produce human-like dialogue, admit mistakes, and reject inappropriate requests. Users leverage the chatbot to execute diverse tasks: generating code snippets, drafting professional emails, summarizing technical documents, and even creating original images via DALL-E integration. It functions as a powerful, multi-purpose tool for rapid content creation and information retrieval.
Compose Email
Loading recent emails...