# The YOLO26 MLX Build Challenge — May 2026

**URL:** <https://community.webai.com/t/the-yolo26-mlx-build-challenge-may-2026/16>\
**Category:** Build Challenges\
**Created:** [May 15, 2026, 12:34am UTC](https://community.webai.com/t/the-yolo26-mlx-build-challenge-may-2026/16 "2026-05-15T00:34:48Z")\
**Posts on this page:** 20\
**Page:** 1

<div class="post-metadata">

**Author:** ![jayatwebai](https://yyz2.discourse-cdn.com/flex056/user_avatar/community.webai.com/jayatwebai/32/6_2.png) [@jayatwebai](https://community.webai.com/u/jayatwebai)\
**Post date:** [May 15, 2026, 12:34am UTC](https://community.webai.com/t/the-yolo26-mlx-build-challenge-may-2026/16/1 "2026-05-15T00:34:48Z")

</div>

## **The YOLO26 MLX Build Challenge**

**Build with our open source YOLO26 MLX model. Seven days. On-device. Show us what you’ve got.**

_Co-hosted by webAI, HackAI, and AITX + Antler_

### **The challenge**

We’re giving you 7 days to build a demo using YOLO26 MLX, our open source, Apple Silicon-native object detection model. Pick a track, pick an idea, ship it.

### **Our co-hosts**

This challenge is co-hosted with two of Austin’s strongest AI builder communities:

**[HackAI](https://luma.com/io3nuky7):** kicking us off on May 18 at Capital Factory during their monthly community session. We’ll be presenting the challenge alongside other Austin AI companies, with open hack and network time after.

**[AITX](https://luma.com/webai-yolo):** running our deeper technical evening on May 19 at [Antler](https://www.antler.co/), where Mitch and Fatih from our AI team walk through how we built YOLO26 MLX. Open Q&A and roundtable after.

Whichever event you attend (or both), you’re in the same challenge with the same pool of builders.

### **Four tracks**

**Useful:** YOLO26 MLX solving a real problem for everyday people.

> _Examples:_ a camera that tells you when your dog leaves the room, a real-time fridge inventory, a posture monitor for desk workers, an object finder for “where are my keys,” a plant health detector.

**Enterprise:** YOLO26 MLX solving a business or industrial problem.

> _Examples:_ tracking foot traffic patterns through a retail floor, defect detection on a manufacturing line, flagging PPE violations without naming individuals, fire or smoke detection on industrial cameras, auditing a warehouse for misplaced inventory.

**Austin-flavored:** YOLO26 MLX solving a problem specific to a place or context in Austin.

> _Examples:_ counting how full the line at Franklin Barbecue is from a webcam feed, watching your front porch for package arrivals, alerting you when the H-E-B parking lot hits capacity, identifying which roommate left the dishes in the sink, tagging the cars in your neighborhood to spot patterns.

**Wild:** YOLO26 MLX doing something intentionally weird, funny, or unexpected.

> _Examples:_ roasting your outfit in real time based on what you’re wearing, narrating your life like a David Attenborough documentary based on the room you’re in, generating AI haikus from objects on your desk, an alarm clock that only stops when it detects you holding a coffee, a Tamagotchi that reacts to your screen.

Build whatever you want inside a track. The examples are just to get you started.

### **Teams**

Build solo or with a team of up to 3 people. Looking for teammates? Drop a note in the [“Who’s in?” topic](https://community.webai.com/t/whos-in-yolo26-mlx-build-challenge/18).

### **What to submit**

1. **Public GitHub repo** with your code. Fork our [starter template](https://github.com/thewebAI/yolo-mlx/blob/main/GUIDE_TRAINING_BENCHMARK.md) to get a standard structure with a working YOLO26 MLX inference script. Forking is strongly recommended but not required.

2. **README that follows this checklist:**

3. **60-second demo video:** screen record or phone-record your demo working. No narration needed, no production polish required. That said, get creative if you want to. A great demo video makes your build memorable.

4. **Social post on X or LinkedIn:** share your demo video on at least one platform, tagged with #YOLOMLX and tagging webAI. Cross-posting on both is encouraged and helps your build get seen by more people.

5. **Confirmation you’ve completed the [registration and acceptance form](https://docs.google.com/forms/d/e/1FAIpQLSfVlgYcREQdsAFRs6Jm_uCz5_BqQYVw5gmCAFkXvVtS-qvSug/viewform?usp=header)** so we know your submission counts.

### **Rules**

- **One submission per individual or team.** Pick your best build, ship it.

- **Pre-existing code is okay as scaffolding.** Your YOLO26 MLX integration and the core build must happen during the challenge window (May 18-24).

- **AI assistants are encouraged.** Cursor, Copilot, Claude, whatever helps you ship. We’re an AI company. Use AI to build.

- **Open source code only.** Don’t include proprietary code from your day job or anywhere else you don’t have rights to publish.

### **Terms & Conditions**

By submitting to this challenge, you agree to the [Terms & Conditions](https://community.webai.com/t/terms-conditions-yolo26-mlx-build-challenge/19/). This covers IP ownership, prize eligibility, and a few other legal basics.

To enter, [fill out the registration and acceptance form](https://docs.google.com/forms/d/e/1FAIpQLSfVlgYcREQdsAFRs6Jm_uCz5_BqQYVw5gmCAFkXvVtS-qvSug/viewform?usp=header). Required to be eligible for prizes. You can do this anytime before the submission deadline.

### **Getting started**

Before you start building, read our [Getting Started with YOLO26 MLX guide](https://community.webai.com/t/getting-started-guide-yolo26-mlx-build-challenge/20). It covers setup, requirements, a working hello-world script, and common gotchas.

### **How you’ll be judged**

- **Use of YOLO26 MLX / On-device execution (20 pts):** Is YOLO26 MLX meaningfully used? Does the demo run locally on Apple Silicon? Is on-device inference central, not decorative?

- **Demo quality / Shipping completeness (20 pts):** Does it actually work live? Is the user flow clear? Is it stable enough to understand the idea without hand-waving?

- **Impact / Usefulness (20 pts):** Does it solve a real problem or create a clearly valuable experience? Is the target user obvious?

- **Technical execution (15 pts):** Is the implementation thoughtful? Good latency, model integration, camera/input handling, architecture, edge-case handling?

- **Creativity / Originality (15 pts):** Is the idea fresh, clever, surprising, or differentiated from obvious object-detection demos?

- **Presentation / Storytelling (10 pts):** Did the team explain the problem, solution, demo, and why it matters clearly?

**Total: 100 points**

**Judges:**

- Mitch DePree, ML Engineer, webAI

- Fatih Altay, ML Engineer, webAI

- Hossein Moghimifam, VP of AI, webAI

- Jay Peredo, Community Lead, webAI

- Sam Avila, ML Engineer, webAI

### **Prizes**

- **Track winners (4):** $1000 per team + featured by webAI across blog and socials. A highlight post on our blog, a thread on X, a LinkedIn post, and amplification across our community channels. Cash prize is awarded to the team and split at the team’s discretion.

- **Every submission:** included in our recap blog post and amplified across webAI socials.

### **Timeline**

- **Mon May 18, 6:00 - 8:00 PM CT:** Kickoff #1 at HackAI (Capital Factory)

- **Tue May 19, 5:30 - 8:00 PM CT:** Kickoff #2 at AITX tech talk (Antler)

- **Thu May 21:** Mid-challenge check-in. Our AI/ML team will be actively answering questions in the [Q&A topic for this challenge](https://community.webai.com/t/q-a-yolo26-mlx-build-challenge/17)

- **Sun May 24, 11:59pm PT:** Submissions due

- **Wed May 27:** Winners announced

### **How to join**

To enter the challenge:

1. **Sign up** for a [community.webai.com](https://community.webai.com/) account if you haven’t already (free, takes 30 seconds)

2. **Register** by filling out the registration and acceptance form (required to be eligible for prizes)

3. **Build** it using YOLO26 MLX, following the requirements above

4. **Submit** by replying directly to this topic by Sun May 24, 11:59pm PT. Include your GitHub repo link, demo video, and social post link in your reply

Submissions posted anywhere else won’t count, so make sure your reply lands here.

The [Getting Started guide](https://community.webai.com/t/getting-started-guide-yolo26-mlx-build-challenge/20) covers setup and a working hello-world. The [Q&A topic for this challenge](https://community.webai.com/t/q-a-yolo26-mlx-build-challenge/17) is where to ask questions. The [“Who’s in?” topic](https://community.webai.com/t/whos-in-yolo26-mlx-build-challenge/18) is where to tell us you’re participating and find teammates.

---

<div class="post-metadata">

**Author:** ![jayatwebai](https://yyz2.discourse-cdn.com/flex056/user_avatar/community.webai.com/jayatwebai/32/6_2.png) [@jayatwebai](https://community.webai.com/u/jayatwebai)\
**Post date:** [May 15, 2026, 2:24pm UTC](https://community.webai.com/t/the-yolo26-mlx-build-challenge-may-2026/16/2 "2026-05-15T14:24:59Z")

</div>



---

<div class="post-metadata">

**Author:** ![jayatwebai](https://yyz2.discourse-cdn.com/flex056/user_avatar/community.webai.com/jayatwebai/32/6_2.png) [@jayatwebai](https://community.webai.com/u/jayatwebai)\
**Post date:** [May 20, 2026, 12:13am UTC](https://community.webai.com/t/the-yolo26-mlx-build-challenge-may-2026/16/3 "2026-05-20T00:13:14Z")

</div>

The starter template hyperlink has been updated, but you can also find it here: [yolo-mlx/GUIDE\_TRAINING\_BENCHMARK.md at main · thewebAI/yolo-mlx · GitHub](https://github.com/thewebAI/yolo-mlx/blob/main/GUIDE_TRAINING_BENCHMARK.md)

---

<div class="post-metadata">

**Author:** ![alex\_kramer](https://yyz2.discourse-cdn.com/flex056/user_avatar/community.webai.com/alex_kramer/32/68_2.png) [@alex\_kramer](https://community.webai.com/u/alex_kramer)\
**Post date:** [May 20, 2026, 7:08pm UTC](https://community.webai.com/t/the-yolo26-mlx-build-challenge-may-2026/16/4 "2026-05-20T19:08:40Z")

</div>

Thanks for updating the Starter Template - I couldn’t find it yesterday! Thanks again!

---

<div class="post-metadata">

**Author:** ![bishopz](https://yyz2.discourse-cdn.com/flex056/user_avatar/community.webai.com/bishopz/32/36_2.png) [@bishopz](https://community.webai.com/u/bishopz)\
**Post date:** [May 22, 2026, 11:35pm UTC](https://community.webai.com/t/the-yolo26-mlx-build-challenge-may-2026/16/5 "2026-05-22T23:35:40Z")

</div>

Yolo Game Part 1  
Repo

> **[GitHub - bishopZ/yolo-game: A scavenger hunt game powered by local-first...](https://github.com/bishopZ/yolo-game)**
>
> A scavenger hunt game powered by local-first segmentation

60 second video

> **[YOLO26 MLX Challenge video](https://framerate.tv/watch/9ca92aa2-7b97-429b-9297-74c8ca519e97)**
>
> This is a sample video I made for the YOLO26 MLX Challenge. It demonstrates the Yolo Game App that I made for the challenge. Details and Download can be found o

---

<div class="post-metadata">

**Author:** ![bishopz](https://yyz2.discourse-cdn.com/flex056/user_avatar/community.webai.com/bishopz/32/36_2.png) [@bishopz](https://community.webai.com/u/bishopz)\
**Post date:** [May 22, 2026, 11:40pm UTC](https://community.webai.com/t/the-yolo26-mlx-build-challenge-may-2026/16/6 "2026-05-22T23:40:57Z")

</div>

Yolo Game part 2

Social post

> **[Yolo Game | Bishop Zareh](https://www.linkedin.com/feed/update/urn:li:activity:7463682689200594944/)**
>
> I made a scavenger hunt game that runs locally on Apple Silicon with no cloud, credits or tokens. Fully private, and an ideal icebreaker for any social gathering.
> 
> Free to play. Download now.
> https://lnkd.in/gGKmhMWJ
> 
> I made this game as my entry to...

More info

> **[Yolo Game](https://bishopz.com/articles/yolo-game)**
>
> A local-first scavenger hunt on Apple Silicon with YOLO26 MLX, no cloud, and a reason to leave your desk.

---

<div class="post-metadata">

**Author:** ![nomad-link](https://avatars.discourse-cdn.com/v4/letter/n/85e7bf/32.png) [@nomad-link](https://community.webai.com/u/nomad-link)\
**Post date:** [May 23, 2026, 10:26pm UTC](https://community.webai.com/t/the-yolo26-mlx-build-challenge-may-2026/16/7 "2026-05-23T22:26:45Z")

</div>

Project: SENTINEL — on-device visual triage AI  
Track: Enterprise  
Team: nomad-link-id + Lexi Armstrong

SENTINEL classifies every person in frame as T1 — immediate (lying down), T2 — delayed (sitting), or T3 — ambulatory (standing) using YOLO26 (yolo26n) + MLX. Built for mass-casualty scenarios where network connectivity is unavailable, compromised, or operationally forbidden: tactical edge, disaster zones, austere medical settings, secured facilities.

The entire pipeline runs on Apple Silicon. The demo video was recorded with WiFi disabled — zero network egress isn’t a marketing claim, it’s the architecture.

Repo: [GitHub - nomad-link-id/sentinel-mlx: On-device visual triage AI · YOLO26 + Apple MLX · Zero network egress · webAI Build Challenge May 2026 · GitHub](https://github.com/nomad-link-id/sentinel-mlx)  
Demo (55s): [https://youtu.be/c2v5Mdg5fpw](https://youtu.be/c2v5Mdg5fpw)

Hardware: MacBook Pro 14-inch (Nov 2024), Apple M4, 32 GB RAM  
Model variant: yolo26n  
Performance: ~16 FPS at 720p, ~45 ms per-frame inference, 0 bytes egress

Single-file Python, single-thread synchronous loop. yolo-mlx 0.3.1, mlx 0.30.6, OpenCV 4.13. Forked the official starter to focus the 7 days on the triage classification logic, the SENTINEL dashboard UI overlay, and end-to-end stability. Architecture rationale in docs/ARCHITECTURE\_DECISIONS.md.

Disclaimer: Research demonstration, not a medical device.

Thanks to Mitch, Fatih, Hossein, Jay, and Sam for the YOLO26-MLX release and the build challenge.

-– nomad-link-id & Lexi

---

<div class="post-metadata">

**Author:** ![Jordaaan](https://yyz2.discourse-cdn.com/flex056/user_avatar/community.webai.com/jordaaan/32/39_2.png) [@Jordaaan](https://community.webai.com/u/Jordaaan)\
**Post date:** [May 24, 2026, 1:54pm UTC](https://community.webai.com/t/the-yolo26-mlx-build-challenge-may-2026/16/8 "2026-05-24T13:54:23Z")

</div>

Project: Meta YOLO buyer  
Track: Austin & Enterprise  
Team: Jordaaan

Repo: [https://github.com/Organized-AI/YOLO-MLX-Hack](https://github.com/Organized-AI/YOLO-MLX-Hack)  
Video and project brief: [Meta YOLO Buyer — YOLO26 MLX Media Buying Harness](https://hack.organizedai.vip/yolo-mlx)

- Built a Cloudflare Workers backend to coordinate the full loop: analyze creative → propose remix → test variant → learn from results.

- Used Durable Objects to manage campaign-level state and prevent conflicting updates across campaign, ad set, ad, and creative workflows.

- Stored YOLO/MLX computer vision outputs as structured inspection signals, so the backend knows what changed: headline load, product cue, subject crop, hook, and CTA.

- Designed the system around human-gated creative changes, with proposed updates blocked from publishing until reviewed.

- Added a 98% confidence threshold before any future automated action is allowed, keeping automation controlled and evidence-based.

- Built to support Agent orchestration (i.e. Hermes can inspect past results and learn how to use this from top to bottom)

---

<div class="post-metadata">

**Author:** ![okigan](https://yyz2.discourse-cdn.com/flex056/user_avatar/community.webai.com/okigan/32/37_2.png) [@okigan](https://community.webai.com/u/okigan)\
**Post date:** [May 24, 2026, 3:10pm UTC](https://community.webai.com/t/the-yolo26-mlx-build-challenge-may-2026/16/9 "2026-05-24T15:10:18Z")

</div>

Project: ScreenSense — Framework for training of Real-Time UI Element Detection  
Track: Useful / Enterprise  
Team: okigan

ScreenSense - establishes framework for training UI elements detection and demostrates on 17 types of UI elements (buttons, checkboxes, text inputs, dropdowns, sliders, toggles, cards, dialogs, etc.) directly from screen pixels using YOLO26-Medium + MLX. Built for computer use agents, UI test automation, accessibility auditing, and RPA — anywhere you need to find interactive controls without DOM access.

The framework constructed as a pipeline: synthetic data generation, training from scratch (not fine-tune), and live inference — runs on a single MacBook. No cloud, no manual labeling, no platform-specific APIs.

Repo: [GitHub - okigan/yolo-screensense · GitHub](https://github.com/okigan/yolo-screensense)  
Demo: [https://youtu.be/5GBi5RSLAc0](https://youtu.be/5GBi5RSLAc0) (short version) [https://youtu.be/rRlg9WzKS3s](https://youtu.be/rRlg9WzKS3s) (longer version)

Hardware: MacBook Pro, Apple M1 Max, 64 GB RAM  
Model variant: yolo26m  
Performance: ~3.5 FPS at 640px, ~288 ms per-frame inference, 0 network calls

Key results: mAP50 0.903, Precision 0.834, Recall 0.947 across 17 GUI element classes. Trained from scratch in ~7.5 hours on 2,000 fully synthetic images generated via Playwright in 10 minutes. Zero manual annotation — bounding boxes extracted from DOM after headless render.

Live demo supports interactive region selection and adjustable confidence threshold on any application window.

Whole framework / pipeline can be retargeted to specific UI framework(s) by adjusting training data generation, auto labeling procedure and augmentation procedure.

Thanks to the webAI team for YOLO26-MLX and the build challenge.

#YOLOMLX

---

<div class="post-metadata">

**Author:** ![nanaagyei](https://avatars.discourse-cdn.com/v4/letter/n/7993a0/32.png) [@nanaagyei](https://community.webai.com/u/nanaagyei)\
**Post date:** [May 24, 2026, 10:30pm UTC](https://community.webai.com/t/the-yolo26-mlx-build-challenge-may-2026/16/10 "2026-05-24T22:30:50Z")

</div>

Name: AccessLens — Real-Time Scene Narrator for the Visually Impaired

Track: Useful

Team: Prince

AccessLens narrates a physical environment in real-time using YOLO26 (yolo26n) + MLX. Point a MacBook camera at a room and it speaks what’s there — object identity, spatial position, proximity — with priority-scored cadence that prevents information overload. Built for blind and low-vision users who need spatial awareness without surrendering visual privacy to a cloud service.

The entire pipeline runs on Apple Silicon. Camera frames never leave the device — not to a server, not to an API, not to a log. Zero network egress is the architecture, not a configuration option.

Repo: [GitHub - nanaagyei/accesslens · GitHub](https://github.com/nanaagyei/accesslens.git)  
Demo (52s): [https://youtu.be/YFThk8kSvag](https://youtu.be/YFThk8kSvag)

Hardware: MacBook Pro, Apple SiliconModel variant: yolo26nPerformance: ~10 FPS capture, ~85ms p50 inference, \<130ms detection-to-speech, 0 bytes egress

Pipeline: Camera → WebSocket (localhost) → YOLO26 MLX → spatial zone classification → priority narrator → Web Speech API. Single-machine, two-process architecture: Python/FastAPI backend for inference, Next.js frontend for capture + narration logic.

Interaction model: Space (describe scene), F (find/search for object by class), M (mute), B (blind mode — screen off, narration continues), ? (shortcuts overlay). Voice search highlights matching bounding boxes and speaks location.

Architecture decisions: Narration logic is a pure function (no DOM, no speech calls) — fully unit tested. Tracker uses frame-to-frame IoU matching with label locking to prevent flickering. WebSocket uses backpressure (drops frames if inference in-flight) rather than queuing. Confidence-scaled bounding box opacity gives visual feedback on detection certainty.

---

<div class="post-metadata">

**Author:** ![tim](https://yyz2.discourse-cdn.com/flex056/user_avatar/community.webai.com/tim/32/18_2.png) [@tim](https://community.webai.com/u/tim)\
**Post date:** [May 24, 2026, 11:36pm UTC](https://community.webai.com/t/the-yolo26-mlx-build-challenge-may-2026/16/11 "2026-05-24T23:36:27Z")

</div>

Hey this is team provenance.guru

github repo:

> **[GitHub - timlefkowitz/GalleryAnalytics](https://github.com/timlefkowitz/GalleryAnalytics)**
>
> Contribute to timlefkowitz/GalleryAnalytics development by creating an account on GitHub.

linkedin post (includes video):

> **[\#yolomlx #art #institutions #artist #antler | Timothy Lefkowitz](https://www.linkedin.com/feed/update/urn:li:activity:7464458781356105728/)**
>
> We entered The YOLO26 MLX Build Challenge — May 2026 
> hackathon at Antler with webAI
> 
> This was an awesome experience with computer vision AND to come out of it with a new product for Provenance.guru is a blessing 
> 
> Having this system setup in...

Thank you team webAi this was fun and will be continuing to work on our new product at provenance.guru 😃

---

<div class="post-metadata">

**Author:** ![Karthik\_Barma](https://yyz2.discourse-cdn.com/flex056/user_avatar/community.webai.com/karthik_barma/32/15_2.png) [@Karthik\_Barma](https://community.webai.com/u/Karthik_Barma)\
**Post date:** [May 25, 2026, 12:06am UTC](https://community.webai.com/t/the-yolo26-mlx-build-challenge-may-2026/16/13 "2026-05-25T00:06:24Z")

</div>

## Sovereign Vision - Enterprise track submission

**Team:** Karthik Barma (solo)

**Project tagline:** The first on-device enterprise vision system whose privacy is enforced by code, not by policy. Seven cryptographically-audited constitutional rules redact every person bounding box, hash every face region, drop every track ID, and add calibrated differential-privacy noise to every aggregate, before any output exists.

* * *

### Submission deliverables

- **GitHub repo:** [GitHub - TheBarmaEffect/sovereign-vision: The first on-device enterprise vision system that is GDPR-compliant by design, not by policy. YOLO26-MLX submission, Enterprise track. Powered by the Glass Box Framework · GitHub](https://github.com/TheBarmaEffect/sovereign-vision)
- **60-second demo video:** [sovereign-vision/assets/demo\_60s.mp4 at main · TheBarmaEffect/sovereign-vision · GitHub](https://github.com/TheBarmaEffect/sovereign-vision/blob/main/assets/demo_60s.mp4)
- **PyPI package (live):** [sovereign-vision · PyPI](https://pypi.org/project/sovereign-vision/)
- **Homebrew tap (live):** [GitHub - TheBarmaEffect/sovereign-vision-tap: Official Homebrew tap for Sovereign Vision - on-device enterprise vision firewall. · GitHub](https://github.com/TheBarmaEffect/sovereign-vision-tap)

### What it does

Intercepts every YOLO26 MLX inference in flight and enforces a 7-rule constitutional firewall. Detections that leave the pipeline are aggregate-only, PII-redacted, and cryptographically attested. Each frame issues a self-attested compliance certificate. Certificates chain into a Merkle tree. Session roots can be anchored to a DigiCert RFC 3161 trusted timestamp so a regulator three years from now can verify exactly when the session happened.

### Why I built it

Most enterprise CV deployments die in legal review. GDPR Article 4 makes a person’s spatial location personal data. Article 9 covers face data. Recital 30 covers track IDs. Every off-the-shelf CV system produces PII the instant it generates a bounding box, which is why most factories, stores, and hospitals still don’t deploy the cameras they already own. Sovereign Vision rebuilds the stack so PII is impossible at the type level.

### How to run

```bash
pip install sovereign-vision
sovereign demo

```

---

<div class="post-metadata">

**Author:** ![nik875](https://avatars.discourse-cdn.com/v4/letter/n/e19b73/32.png) [@nik875](https://community.webai.com/u/nik875)\
**Post date:** [May 25, 2026, 12:15am UTC](https://community.webai.com/t/the-yolo26-mlx-build-challenge-may-2026/16/14 "2026-05-25T00:15:46Z")

</div>

Name: Daedalus-RT, Undetectably Flood your Hackathon Competitors with False Positive Detections!

Track: Weird/wacky

Team: Nikhil Kalidasu (solo)

Wouldn’t it be funny to mess with literally every other submission to the YOLO26 MLX Build Challenge? Introducing Daedalus-RT, a real-time, on-device adversarial attack that floods YOLO26-n with high-confidence false-positive detection boxes. With effectively zero added latency, compositing a single adversarial filter onto every video frame can completely destroy the model’s ability to detect real objects. Detailed attack breakdown is in the repo’s README.

Repo: [GitHub - nik875/Daedalus-RT: Real-time version of 'Daedalus: Breaking Non-Maximum Suppression in Object Detection via Adversarial Examples' via universal adversarial perturbation, tuned for YOLO26 · GitHub](https://github.com/nik875/Daedalus-RT)  
Demo: [https://youtube.com/shorts/gmYOZACKt3Q?feature=share](https://youtube.com/shorts/gmYOZACKt3Q?feature=share)  
LinkedIn Post: [#yolomlx | Nikhil Kalidasu](https://www.linkedin.com/feed/update/urn:li:activity:7464466877252009984/)

Inference hardware: Apple M1 Pro, 16GB RAM  
Model variant: YOLO26-n  
Latency: 50+ FPS baseline, undetectable latency change under attack.

**Why did I build this?**

Because it’s fun! But also because I want a job 👉 👈

I would love to work with webAI on developing similar low-latency on-device AI, and hope this 1-day sprint project showcases my technical depth. I applied last week for the AI Research Scientist role. I would greatly appreciate an interview, I pinky promise I know what I’m talking about 🙏 !

---

<div class="post-metadata">

**Author:** ![Saksham\_Adhikari](https://yyz2.discourse-cdn.com/flex056/user_avatar/community.webai.com/saksham_adhikari/32/26_2.png) [@Saksham\_Adhikari](https://community.webai.com/u/Saksham_Adhikari)\
**Post date:** [May 25, 2026, 1:21am UTC](https://community.webai.com/t/the-yolo26-mlx-build-challenge-may-2026/16/15 "2026-05-25T01:21:42Z")

</div>

Meet Don’t Wake Up - the alarm clock that does not trust you.

Track: Wild

Team: Quamos (Saksham Adhikari and Kusum Bhattarai Sharma)

Why settle for a normal alarm clock when you can get an alarm clock which rage baits you out of bed?  
That is what we built.

**Don’t Wake Up** is an MLX-powered computer vision alarm clock that refuses to shut up until you prove you are actually awake. Not by tapping a button. Not by shaking your phone. Not by mumbling at Siri. You have to get out of bed, stand in front of the camera, do the 6-7 movement ( [Reddit - Please wait for verification](https://www.reddit.com/r/Teachers/comments/1p13onp/can_someone_please_explain_what_67_means/) ), and then show a real coffee mug to turn it off.

If you try to cheat by showing a picture of a mug on your phone, it calls you out: “na na buddy, that won’t work.”

The alarm only stops when the full wake-up ritual is completed.  
Here is the demo video :

[![](https://canada1.discourse-cdn.com/flex056/uploads/webai/original/1X/39418b0dddb56f16be198f6ae075ad200120a3cd.jpeg "YOLO26 Build Challenge: Team Quamos demo video") ](https://www.youtube.com/watch?v=vi7vPIOT0Yw)

GitHub Repo: [GitHub - Tar-ive/6\_7: A YOLO alarm system · GitHub](https://github.com/Tar-ive/6_7)

X post: [https://x.com/saksham\_adh/status/2058712860935512244](https://x.com/saksham_adh/status/2058712860935512244)

The hardest part of this project was that YOLO26 pose support was not fully available in the MLX YOLO stack yet. So we patched it and opened a PR so we could complete the project. [Add YOLO26 pose inference support by Tar-ive · Pull Request #6 · thewebAI/yolo-mlx · GitHub](https://github.com/thewebAI/yolo-mlx/pull/6)

Thanks to the WebAI team for the Challenge, we really enjoyed building with the MLX-webai stack.

---

<div class="post-metadata">

**Author:** ![KVC](https://yyz2.discourse-cdn.com/flex056/user_avatar/community.webai.com/kvc/32/65_2.png) [@KVC](https://community.webai.com/u/KVC)\
**Post date:** [May 25, 2026, 1:49am UTC](https://community.webai.com/t/the-yolo26-mlx-build-challenge-may-2026/16/16 "2026-05-25T01:49:50Z")

</div>

Category: Useful

I built \*\ ***Vault** \*\* for webAI’s YOLO26-MLX build challenge — a privacy-first home inventory app that uses YOLO26 fine-tuned on household objects to catalog items for insurance documentation, entirely on-device on iPhone.

The technical bet I’m most proud of: \*\ ***RoomPlan LiDAR and YOLO26 object detection run simultaneously on a single ARSession.** \*\* One camera feed powers both. RoomPlan builds the 3D mesh; YOLO inference fires at 5 fps on the same `arSession.currentFrame` stream — without stealing the ARSessionDelegate, so RoomPlan’s pipeline stays intact. Two real-time ML pipelines, one frame source, both running on the iPhone’s Apple GPU.

The training pipeline:

-\> Stock yolo26-s ported by hand to MLX-Swift for iOS (~38 MB, MAE=0 parity with PyTorch)

-\> Fine-tuned yolo26-m on a merged dataset of HomeObjects-3K + Roboflow household (3,622 images, 32 classes covering furniture and small appliances)

-\> Trained for 50 epochs in webAI’s `yolo-mlx` on an M1 Max, batch 8, ~7 hours. \*\ ***Zero PyTorch dependency at runtime.** \*\*

-\> Best checkpoint at epoch 45: \*\ ***mAP50 = 0.744, mAP50-95 = 0.57** \*\*

-\> The resulting 84 MB `.safetensors` drops directly into the iOS app bundle

End-to-end MLX. Training stays on the user’s Mac. Inference stays on the user’s phone. Photos never leave the device — only the structured catalog (names, counts, user-typed values) goes to an LLM for the priced report.

This is the on-device AI story that webAI’s MLX stack actually enables: a complete, vertical app where the model can be customized by the developer locally and shipped to users without ever touching a cloud GPU.

Deck: 🖼 [https://slideshow-kohl.vercel.app](https://slideshow-kohl.vercel.app) includes the 1 min video

Source: 🔗 [GitHub - PromptForcePrime/vault-ext: Vault is iPhone all that allows users to privately capture and catalog their home assets for insurance purposes without seeing photos over to any cloud services. · GitHub](https://github.com/PromptForcePrime/vault-ext)

-– @W2OPQR

---

<div class="post-metadata">

**Author:** ![KVC](https://yyz2.discourse-cdn.com/flex056/user_avatar/community.webai.com/kvc/32/65_2.png) [@KVC](https://community.webai.com/u/KVC)\
**Post date:** [May 25, 2026, 1:50am UTC](https://community.webai.com/t/the-yolo26-mlx-build-challenge-may-2026/16/17 "2026-05-25T01:50:21Z")

</div>

Demo: 🎥 [https://youtube.com/shorts/9kwbLM0gVxg](https://youtube.com/shorts/9kwbLM0gVxg)

---

<div class="post-metadata">

**Author:** ![vel](https://avatars.discourse-cdn.com/v4/letter/v/eb9ed0/32.png) [@vel](https://community.webai.com/u/vel)\
**Post date:** [May 25, 2026, 2:09am UTC](https://community.webai.com/t/the-yolo26-mlx-build-challenge-may-2026/16/18 "2026-05-25T02:09:45Z")

</div>

CareSight is a local-first caregiver awareness prototype that watches for possible care events, such as a floor stay or a missing loved one, stores the event locally, and turns it into a human-readable review story.

Track: **Enterprise**

Why Enterprise: the demo runs in a home-style setup, but the product path is caregiver operations: care teams, assisted-living workflows, local audit trails, and human-reviewed handoff queues. The current build is immediately useful for family caregiver awareness, while the same bounded loop can scale toward enterprise care operations with the right deployment, compliance, and workflow layers.

Each review answers:

- What was likely observed
- Where it happened
- What evidence exists
- What follow-up may be needed

The bounded agentic stack:

- YOLO26 MLX for local perception
- SQLite as the local black-box event record
- Gemma-family MLX model for caregiver-facing drafts
- Hermes Agent for staged, no-send handoff workflows
- Holler 0.6B 6-bit TTS with Dakota voice for approved local readbacks
- OBS / FaceTime as optional, human-approved handoff surfaces

The key constraint: agents assist the loop, but humans remain the authority. No medical-device claim, no autonomous emergency dispatch, and raw camera context stays local by default.

Social Media Posts - All have Video:

- LinkedIn: [Here](https://www.linkedin.com/posts/spajewski_yolomlx-ugcPost-7464486100519075840-ZT_G)
- Twitter: [Here](https://x.com/Velcrafting/status/2058712618701836308)
- Tiktok: [Here](https://www.tiktok.com/t/ZP8pt8Cre/)
- Youtube: [Here](https://youtube.com/shorts/wGKUoL215SI?feature=share)

Github

- Main Repo: [Here](https://github.com/Vel-Labs/CareSight)
- Hackathon Docs: [Here](https://github.com/Vel-Labs/CareSight/tree/main/hackathon)
- Side Project During Hackathon: [Here](https://github.com/Vel-Labs/yolo26-mlx-swift)
  - Swift implementation for YOLO26 MLX for Mobile apps / usage

---

<div class="post-metadata">

**Author:** ![kayliejayy](https://avatars.discourse-cdn.com/v4/letter/k/ecccb3/32.png) [@kayliejayy](https://community.webai.com/u/kayliejayy)\
**Post date:** [May 25, 2026, 2:34am UTC](https://community.webai.com/t/the-yolo26-mlx-build-challenge-may-2026/16/19 "2026-05-25T02:34:20Z")

</div>

> [@jayatwebai](#):
>
> 1. **Submit** by replying directly to this topic by Sun May 24, 11:59pm PT. Include your GitHub repo link, demo video, and social post link in your reply
> 
> Submissions posted anywhere else won’t count, so make sure your reply lands here.

Hi everyone! Here is my submission for LiftLens.

GitHub repo: [Here](https://github.com/kayliejayy/LiftLens)

Demo video / social post:

- X: [https://x.com/kaylie\_jayy/status/2058409579793265140?s=20](https://x.com/kaylie_jayy/status/2058409579793265140?s=20)
- TikTok: [TikTok - Make Your Day](https://www.tiktok.com/@jordyn.2.0/video/7643485233811115277)
- LinkedIn: [#yolomlx | Kaylie Rader](https://www.linkedin.com/posts/kaylierader_yolomlx-ugcPost-7464341962574225409-Ggt6/)

Track: **Useful**

LiftLens is an iOS fitness app prototype built to help beginners and returning gym users feel more confident starting a workout. The core idea is simple: do the workout, and LiftLens logs it.

Instead of expecting users to already know what to do, remember every set, or manually track everything afterward, LiftLens guides them through onboarding, recommends a beginner-friendly workout, uses the camera to track a squat set, counts reps, and saves a local workout summary.

The value add is reducing the intimidation and mental load of getting started at the gym. LiftLens focuses on confidence, clarity, and repeatability: helping users understand what to do, complete a set, and walk away with a record of their workout.

YOLO26n MLX is used as the on-device person-tracking layer, with Swift integration through Vel-Labs/yolo26-mlx-swift, while the app layers workout flow, rep tracking, and local summaries on top.

---

<div class="post-metadata">

**Author:** ![sariknanaki](https://avatars.discourse-cdn.com/v4/letter/s/5f9b8f/32.png) [@sariknanaki](https://community.webai.com/u/sariknanaki)\
**Post date:** [May 25, 2026, 2:59am UTC](https://community.webai.com/t/the-yolo26-mlx-build-challenge-may-2026/16/20 "2026-05-25T02:59:41Z")

</div>

Hello fellow button-pressers!

**Reality Firewall** — a local physical-world policy engine for Apple Silicon. A webcam streams into **YOLO26 MLX** (`yolo26n.npz`, MLX/Metal), and a Python policy engine evaluates each frame against JSON rule packs, emitting `ALLOW` / `WARN` / `DENY` / `APPROVAL_REQUIRED` / `UNLOCKED` decisions with SHA-256 chained audit receipts.

**Stack**

- **Vision:** YOLO26 MLX via `from yolo26mlx import YOLO`, `model.predict(frame, conf=0.25)`; ~real-time on M-series.
- **Hands:** MediaPipe HandLandmarker (index fingertip, EMA-smoothed) for point-to-unlock.
- **Backend:** FastAPI + WebSocket vision stream, plus an HTTP `/api/select-detection` endpoint so unlock confirms are not blocked by inference.
- **Policy engine:** Rule packs (`home.json`, `enterprise.json`) with conditions like `liquid_near_laptop`, `overlap_zone`, `inventory_mismatch`, `ppe_style_proxy`, `after_hours_person`, with stable/clear hysteresis to prevent event spam.
- **Audit:** Canonical JSON SHA-256 receipts chained per event; JSONL export.
- **Frontend:** React + Vite UI with detection overlays, decision card, event timeline, and a linked receipt block view.

**Reality Passphrase.** Point at **cup → tv/monitor → dog** with your index finger; ~600ms dwell per target unlocks Enterprise Mode and swaps the active policy pack.

Repo: [GitHub - sariknanaki/yolo26-hackathon · GitHub](https://github.com/sariknanaki/yolo26-hackathon)

Demo: [https://www.youtube.com/watch?v=WY9C\_8AhEko](https://www.youtube.com/watch?v=WY9C_8AhEko)

Social Post: [#yolomlx | Kristian Magda](https://www.linkedin.com/feed/update/urn:li:activity:7464504251290132480/)

---

<div class="post-metadata">

**Author:** ![CupOfGeo](https://avatars.discourse-cdn.com/v4/letter/c/a8b319/32.png) [@CupOfGeo](https://community.webai.com/u/CupOfGeo)\
**Post date:** [May 25, 2026, 4:54am UTC](https://community.webai.com/t/the-yolo26-mlx-build-challenge-may-2026/16/21 "2026-05-25T04:54:45Z")

</div>

1. [GitHub - CupOfGeo/yolo-mlx-build-geo: My submission to yolo-mlx build challenge · GitHub](https://github.com/CupOfGeo/yolo-mlx-build-geo)

2. **README that follows this checklist:**

3. [https://www.youtube.com/watch?v=rp6ZqeuvXZQ](https://www.youtube.com/watch?v=rp6ZqeuvXZQ) walk through

4. My LinkedIn post linkedin .com /posts/george-mazzeo\_yolomlx-submission-share-7464535730841755648-kkuz (sorry getting hit with too many links for new account remove the two spaces infront and after the .com)

[Next page](https://community.webai.com/t/the-yolo26-mlx-build-challenge-may-2026/16.md?page=2)
