# Computer Vision Projects at AI Tinkerers

> Canonical HTML: https://aitinkerers.org/technologies/computer-vision
> Markdown URL: https://aitinkerers.org/technologies/computer-vision.md
> Technology record last updated: 2026-02-23T06:12:54Z
> Generated: 2026-09-22T12:37:40Z

Computer Vision (CV) uses deep learning and neural networks (e.g., CNNs) to enable machines to interpret visual data: identifying objects, people, and patterns with high accuracy.

CV is a core AI subfield that actively replicates human sight, processing raw images and video via deep learning models to extract meaningful, actionable data. It is critical for major industries: powering object detection in autonomous vehicles, enabling defect detection in manufacturing quality control, and assisting with diagnostics in medical imaging. This technology drives significant operational efficiency and safety across sectors, with the overall market expected to reach $58.29 billion by 2030 (Grand View Research). We leverage CV to automate tasks that traditionally required human visual inspection.

- Official technology site: https://azure.microsoft.com/en-us/resources/cloud-computing-dictionary/what-is-computer-vision/
- Public AI Tinkerers demos and talks: 22
- Result page: 1 of 1

## Recent Public Talks and Demos

### [Long](https://hong-kong.aitinkerers.org/talks/rsvp_gS_HYkNSquI)

With spec-driven AI development going mainstream and even Garry Tan making his own coding framework, I decided to build a framework myself with ALM (Application Lifecycle Management), Compliance as Code and AI Agents.

- Event context: AI Tinkerers Hong Kong &amp; GBA: AI Demo Night @ Google (May) — 2026-05-07 — Hong Kong
- Public talk page: https://hong-kong.aitinkerers.org/talks/rsvp_gS_HYkNSquI

### [Exercise Posture Tracking Using Computer Vision](https://raleigh.aitinkerers.org/talks/rsvp_QACg5zj28MM)

I've created an unsupervised computer vision algorithm to detect back rounding during deadlift. In short, I've demonstrated that I can measure the curvature of the back in an unsupervised way, and that a rounded back produces a different signal from a flat back. I've used off-the-shelf pose tracking and segmentation models to measure the back curvature. This could help guide learning of proper form during deadlift and other exercises.

- Event context: AI Tinkerers Raleigh Meetup — February 11, 2026 — 2026-02-11 — Raleigh
- Public talk page: https://raleigh.aitinkerers.org/talks/rsvp_QACg5zj28MM

### [Mapping out the Peripheral Nervous System for Neuromodulation Protocols to Prevent/Treat T2D](https://boston.aitinkerers.org/talks/rsvp_xmVxbpf-UR8)

UltraNeuro is a healthcare technology company developing non-invasive solutions for early detection and management of type 2 diabetes (T2D). Its platform includes: • AI-driven facial analysis designed to identify early indicators of gestational and type 2 diabetes risk. • A non-invasive neuromodulation device aimed at improving metabolic function and supporting healthy neural activity associated with glucose regulation.

- Event context: AI Tinkerers Boston Meetup October 2025 — 2025-10-24 — Boston
- Public talk page: https://boston.aitinkerers.org/talks/rsvp_xmVxbpf-UR8

### [FINDER AFRIC](https://nairobi.aitinkerers.org/talks/rsvp_EJjkc96NLNs)

Finder is an AI-powered shopping tool that lets buyers find exactly what they’re looking for just by using images. We’re building a platform where vendors don’t need to be known or followed to make a sale—if they have the product, we’ll match them with the buyer using visual search. I’ll be showcasing how Finder helps small vendors increase visibility and reach, and how we use cutting-edge image recognition to simplify product discovery.

- Event context: AI Tinkerers - Nairobi Inaugural Meetup (April) — 2025-04-09 — Nairobi
- Public talk page: https://nairobi.aitinkerers.org/talks/rsvp_EJjkc96NLNs

### [Self-reflection in agentic workflows](https://zurich.aitinkerers.org/talks/rsvp_4jGx0MOTPXE)

My demo will explore self-reflection in agentic workflows. You will learn how structured self-reflection enhances agentic process quality and prevents system deviations. I will act as an agentic system that executes actions and an AI verifier will monitor my actions and assess my execution. Perhaps I will even be able to trick the verifier ;) Some more details about the demo: I will be impersonating a browser interaction agent (similar to Google's Project Mariner or OpenAI's Operator) and the verifier (based on an LLM with vision) will use natural language description of my actions and screenshots of the webpage as input for verification.

- Event context: AI Tinkerers Zurich - February 6 — 2025-02-06 — Zürich
- Public talk page: https://zurich.aitinkerers.org/talks/rsvp_4jGx0MOTPXE

### [AI Fashion assistant](https://paris.aitinkerers.org/talks/rsvp_a18BJzmpnyY)

AI fashion personal shopper for flexible contextual search and product recommendation

- Event context: AI Tinkerers - Paris Meetup on December 10th — 2024-12-10 — Paris
- Public talk page: https://paris.aitinkerers.org/talks/rsvp_a18BJzmpnyY

### [DeepX Hub](https://palo-alto.aitinkerers.org/talks/rsvp_Ek9BbtrQstI)

Applied AI and Computer Vision platform

- Event context: AI Tinkerers - Palo Alto - November 2024 Meetup — 2024-11-21 — Palo Alto
- Public talk page: https://palo-alto.aitinkerers.org/talks/rsvp_Ek9BbtrQstI

### [Prototyping with ML-powered Livestreaming Apps](https://montreal.aitinkerers.org/talks/rsvp_ShPGnVujPmA)

At Kanastruk, we prototype real-time video and audio processing applications that integrate machine learning and computer vision agents. This presentation will show how you too can incorporate LiveKit’s open-source APIs to quickly develop functional video streaming apps, making it easier to explore and iterate on new ideas in areas like video conferencing and live-streaming.

- Event context: AI Tinkerers - Montreal Inaugural Meetup (October) — 2024-10-29 — Montreal
- Public talk page: https://montreal.aitinkerers.org/talks/rsvp_ShPGnVujPmA

### [DailySnap: Transforming Daily Life into Visual Narratives with AI](https://dubai.aitinkerers.org/talks/rsvp_JkuX3eXbSvs)

In this talk, I will present DailySnap, an innovative AI-driven project that automates the process of capturing daily moments and turning them into visually appealing narratives. The project integrates YOLOv10 for detecting key objects and moments in video data, Llama-3.2-11B-Vision-Instruct for generating descriptive text, and Qwen2.5-Coder-7B-Instruct to create infographics based on these descriptions. Through a live demo, I will walk the audience through the full process, from object detection and data extraction to the final visual output. The session will showcase how multiple AI models can work together to create a seamless automated journaling experience.

- Event context: AI Tinkerers - Dubai Meetup #2 (October) — 2024-10-05 — Dubai
- Public talk page: https://dubai.aitinkerers.org/talks/rsvp_JkuX3eXbSvs

### [Computer Vision and Nanotechnology to Power Immersive Content](https://nyc.aitinkerers.org/talks/rsvp_hAb7erkjEQU)

Leia Inc. utilizes computer vision algorithms and nanotechnology to create immersive content on any device. In this short presentation you will learn how Leia is working with creators and OEMs to bring this exciting future to reality.

- Event context: AI Tinkerers September Meetup at Betaworks — 2024-09-18 — New York City
- Public talk page: https://nyc.aitinkerers.org/talks/rsvp_hAb7erkjEQU

### [Computer Vision.](https://medellin.aitinkerers.org/talks/rsvp_6sn1IkNQMw0)

Entrenamiento e Inferencia con Tensorflow de modelo de segmentación semántica multiproposito usando Databricks.

- Event context: AI Tinkerers Medellín #5 - DataKnow - 28 de Agosto — 2024-08-28 — Medellín
- Public talk page: https://medellin.aitinkerers.org/talks/rsvp_6sn1IkNQMw0

### [Bombi](https://bengaluru.aitinkerers.org/talks/rsvp_PgMIFNUp4Dw)

AR Game to help preschoolers go on adventures with their toys. Kids take their toys’ photos, convert them to living characters and help them navigate adventures through voice/gestures.

- Event context: AI Tinkerers Bangalore - August - RSVP REQUIRED — 2024-08-08 — Bengaluru
- Public talk page: https://bengaluru.aitinkerers.org/talks/rsvp_PgMIFNUp4Dw

### [Building a Sign Language Translator in 5 Minutes](https://sf.aitinkerers.org/talks/rsvp_13YKmPbjwy8)

I'll be building an American Sign Language (ASL) translator using Computer Vision in less than 5 minutes using Roboflow. To do this, I'll be leveraging Roboflow Edge, which allows users to set up an NVIDIA Jetson from scratch using one command. After setting up the Jetson, I'll then modify the logic running on the device using Workflows, a new no-code tool that Roboflow has been building. Once I've modified the logic running on the device, I'll lastly connect the Jetson to a speaker and it will start speaking the sign language gestures that I'm making out loud.

- Event context: AI Tinkerers - San Francisco - Summer Edition - July 2024 — 2024-07-12 — San Francisco
- Public talk page: https://sf.aitinkerers.org/talks/rsvp_13YKmPbjwy8

### [Trabuli beauty](https://bengaluru.aitinkerers.org/talks/rsvp_7eNtwHLiMzM)

Building inspiration led personalized customer journeys for cosmetics brands Demo 1: Plugin for personalizing customer journeys https://youtu.be/mkCMqF1Ayt4 Demo 2: Prototype to create inspirations with makeup look images https://youtube.com/shorts/iXKDAV5Pc_E

- Event context: AI Tinkerers - Bangalore Inaugural - RSVP REQUIRED — 2024-06-02 — Bengaluru
- Public talk page: https://bengaluru.aitinkerers.org/talks/rsvp_7eNtwHLiMzM

### [Chess Predict](https://la.aitinkerers.org/talks/rsvp_HlHrw7-XCZc)

Chesspredict.com is a website for predicting the best move from a screenshot of a digital chess game and providing analysis of it. Plan is to extend to real chessboard photos.

- Event context: May 21st - LA AI Tinkerers Meetup &amp; Demos — 2024-05-22 — Los Angeles
- Public talk page: https://la.aitinkerers.org/talks/rsvp_HlHrw7-XCZc

### [Snoop Hawk](https://seattle.aitinkerers.org/talks/rsvp_JU5-Tehzvac)

Snoop Hawk lets you automate web research. Think visual unit tests. ;) You can set up jobs on any website and have an AI extract insights from it. What's really powerful is that you can schedule these jobs, and soon, set off triggers to notify you of changes. Think design reviews after code changes, competitor analysis, or automated monitoring.

- Event context: AI Tinkerers Seattle - May 2024 — 2024-05-21 — Seattle
- Public talk page: https://seattle.aitinkerers.org/talks/rsvp_JU5-Tehzvac

### [Moondream - vision language model that runs locally in near real-time](https://palo-alto.aitinkerers.org/talks/rsvp__xNyIfVwG0U)

Moondream is a small vision language model that can easily be run locally, but still performs on par or better than many bigger VLMs.

- Event context: AI Tinkerers Palo Alto - Inaugural Meetup — 2024-05-01 — Palo Alto
- Public talk page: https://palo-alto.aitinkerers.org/talks/rsvp__xNyIfVwG0U

### [Gamified Reality](https://sf.aitinkerers.org/talks/rsvp_1ok-_RTnhL4)

This demo shows how you can use the latest small multimodal models to get real-time inferences and then connect those to actions. In this case, I use it to create a gamified reality app where you earn points as you do various real world activities and those are understood by the AI and matched to achievements.

- Event context: AI Tinkerers - San Francisco - April 2024 Meetup — 2024-04-30 — San Francisco
- Public talk page: https://sf.aitinkerers.org/talks/rsvp_1ok-_RTnhL4

### [Document AI Agent](https://berlin.aitinkerers.org/talks/rsvp_rIe52C6q2Yk)

Forto moves thousands of containers each year from Asia to Europe and globally. To do this effectively and as per regulations, we need to handle huge volumes of documents. Every shipment generates hundreds of documents and we need to capture, validate and update data points from these documents. Currently, not only in Forto but globally this is being done manually. Forto is developing an AI based solution to automate this entire process. Our system once implemented will be capable of capturing, validating and updating data from all these documents with limited or no human intervention.

- Event context: AI Tinkerers Berlin - March 21 — 2024-03-21 — Berlin
- Public talk page: https://berlin.aitinkerers.org/talks/rsvp_rIe52C6q2Yk

### [—headful](https://berlin.aitinkerers.org/talks/rsvp_NAh4Rk2Wi1Y)

—headful („dash dash headful“) Pair a web browser with a multimodal model, i.e. give it eyes and make it „—headful“, the opposite of „chrome —headless“. This will be a browser that detects website UI elements and navigates them.

- Event context: AI Tinkerers Berlin - November 24 — 2023-11-24 — Berlin
- Public talk page: https://berlin.aitinkerers.org/talks/rsvp_NAh4Rk2Wi1Y

### [Multimodal Mobile Reasoning](https://sf.aitinkerers.org/talks/rsvp_zszAIf3pvzc)

Hands free mobile usecases for multimodal modals.

- Event context: 🤖🔄🧠 AI Tinkerers SF - October Meetup — 2023-10-26 — San Francisco
- Public talk page: https://sf.aitinkerers.org/talks/rsvp_zszAIf3pvzc

### [Kaizen Copilot](https://seattle.aitinkerers.org/talks/rsvp_JMAd5NagDRI)

Kaizen Copilot helps Industrial Engineers and Production Supervisors improve manufacturing assembly lines using Generative AI and computer vision.

- Event context: AI Tinkerers Seattle - August Meetup — 2023-08-09 — Seattle
- Public talk page: https://seattle.aitinkerers.org/talks/rsvp_JMAd5NagDRI

## Related Technologies

- [Generative AI](https://aitinkerers.org/technologies/generative-ai) ([Markdown](https://aitinkerers.org/technologies/generative-ai.md)) — 45 public demos
- [PyTorch](https://aitinkerers.org/technologies/pytorch) ([Markdown](https://aitinkerers.org/technologies/pytorch.md)) — 273 public demos
- [TensorFlow](https://aitinkerers.org/technologies/tensorflow) ([Markdown](https://aitinkerers.org/technologies/tensorflow.md)) — 90 public demos
- [BERT](https://aitinkerers.org/technologies/bert) ([Markdown](https://aitinkerers.org/technologies/bert.md)) — 179 public demos
- [GPT-3](https://aitinkerers.org/technologies/gpt-3) ([Markdown](https://aitinkerers.org/technologies/gpt-3.md)) — 191 public demos
- [GPT-4](https://aitinkerers.org/technologies/gpt-4) ([Markdown](https://aitinkerers.org/technologies/gpt-4.md)) — 529 public demos
- [Keras](https://aitinkerers.org/technologies/keras) ([Markdown](https://aitinkerers.org/technologies/keras.md)) — 74 public demos
- [ONNX](https://aitinkerers.org/technologies/onnx) ([Markdown](https://aitinkerers.org/technologies/onnx.md)) — 83 public demos
- [scikit-learn](https://aitinkerers.org/technologies/scikit-learn) ([Markdown](https://aitinkerers.org/technologies/scikit-learn.md)) — 84 public demos
- [Object Detection](https://aitinkerers.org/technologies/object-detection) ([Markdown](https://aitinkerers.org/technologies/object-detection.md)) — 2 public demos
- [AI](https://aitinkerers.org/technologies/ai) ([Markdown](https://aitinkerers.org/technologies/ai.md)) — 55 public demos
- [Amazon Personalize](https://aitinkerers.org/technologies/amazon-personalize) ([Markdown](https://aitinkerers.org/technologies/amazon-personalize.md)) — 1 public demo
- [Apache Spark MLlib](https://aitinkerers.org/technologies/apache-spark-mllib) ([Markdown](https://aitinkerers.org/technologies/apache-spark-mllib.md)) — 1 public demo
- [Audio Interface](https://aitinkerers.org/technologies/audio-interface) ([Markdown](https://aitinkerers.org/technologies/audio-interface.md)) — 1 public demo
- [Augmented Reality](https://aitinkerers.org/technologies/augmented-reality) ([Markdown](https://aitinkerers.org/technologies/augmented-reality.md)) — 2 public demos
- [Autoencoder](https://aitinkerers.org/technologies/autoencoder) ([Markdown](https://aitinkerers.org/technologies/autoencoder.md)) — 1 public demo
- [Automation](https://aitinkerers.org/technologies/automation) ([Markdown](https://aitinkerers.org/technologies/automation.md)) — 3 public demos
- [Contextual Search](https://aitinkerers.org/technologies/contextual-search) ([Markdown](https://aitinkerers.org/technologies/contextual-search.md)) — 1 public demo
