Manual exploration
QA engineers repeatedly navigate identical flows across builds.
PocketQA autonomously explores Android apps on real smartphones, finds crashes, visual bugs and performance issues, generates reproducible reports, and verifies fixes by replaying the original journey.
Built by Team OnCall Engineers
Testing
com.demo.shop
Objective
Break the checkout flow
Current action
Changing shipping address
✓ Product discovered
✓ SAVE20 applied
● Changing shipping address
○ Awaiting UI response
LOCAL MODEL ACTIVE
The problem
QA engineers repeatedly navigate identical flows across builds.
Checkout crashed rarely tells a developer what happened immediately before failure.
Developers often spend significant time reproducing a problem before debugging starts.
After every fix, someone must manually verify that the original issue disappeared.
Developers should spend their time fixing bugs, not rediscovering them.
The solution
Before
Manual work at almost every step
After
One continuous automated testing loop
Interactive fake demo
Testing
com.demo.shop
Current action
Changing shipping address
POCKETQA AGENT
IDLE
Test checkout. Apply SAVE20, change the delivery address, and try to complete the purchase.
The ROI
Illustrative example based on configurable assumptions.
Configurable assumptions
Adjust the inputs to model your workflow.
Recovery projection
One bug economics
125 MINUTES
40 MINUTES
Example human time reduction: 85 minutes per bug
Illustrative workflow example, not a measured production benchmark.
Device telemetry
PocketQA correlates failures with actual device behavior on real Android hardware.
CPU
43%
RAM
742 MB
FPS
58
TEMP
39.4°C
BATTERY
3.2% / hr
NETWORK
82 ms
Local AI
LOCAL AI STATUS
Model
On-device VLM
Inference
LOCAL
Cloud requests
0
Screen data uploaded
0 bytes
Status
ACTIVE
PocketQA is designed around local and open-source models for screen understanding and reasoning. Sensitive application context can remain on the testing device.
Architecture
Office Kit can bridge phone and developer environments during the hackathon workflow.
Capabilities
Describe what should work instead of programming every interaction.
AI determines which screens and paths should be tested.
Understands interfaces semantically beyond selectors.
Observes actual performance, thermal, and resource behavior.
Preserves the journey that caused a failure.
Connects runtime behavior with relevant source context.
Turns discovered failures into future tests.
Replays the original failing journey after a fix.
Previous work
PocketQA combines ideas our team has explored across local AI, autonomous agents, interaction replay, and production software.
Local Multimodal AI
Fully local voice-first AI assistant that understands the user’s screen using local vision-language models.
View GitHubInteraction Recording and Replay
Visual browser automation platform that records interactions and converts them into repeatable workflows.
View GitHubAutonomous Voice Agents
AI agent that interacts with customer service systems on behalf of users.
View GitHubThe builders
We build AI systems that interact with the real world, not just chat interfaces.
Builder 01
University of London · Computer Science, AI/ML
Studying AI and Machine Learning at the University of London on a performance scholarship. Shubham has programmed since age 10 and sold his first product at 12. Since then, he has contracted independently and worked as a software engineer and founding engineer for several companies in the US, Dubai, and the UK.
Builder 02
Saptagiri NPS College
Content creator who has collaborated with 40+ brands, bringing a strong audience and product storytelling perspective to the team.
FOCUS
Hackathon achievements
6×
National Hackathon Winners
A track record of turning ambitious ideas into working products and strong live demos.
2026
Top 10
HackBLR, Bengaluru
2025
Winner
CodeClash 2.0 at Google
2025
Top 10
Code for Bharat at Microsoft · 3000+ teams
2025
3rd Prize and Top 10
Code for WIE 3.0, MSIT Delhi · 300+ teams
2024
Top 15
Code the Cubicle 3.0 at Mastercard · 3500+ teams
2024
2nd Place
Live the Code 3.0 · GeeksForGeeks
2024
Top 10 and Special Mention
InnovateX · Delhi Technological University
2024
Best Rookie Hack
Tech Ingenium 5.0 · Ahmedabad University
Why us
We have repeatedly built, demonstrated, and placed at national-level hackathons.
Interaction replay, local VLMs, AI agents, and automation are areas we have already built in.
Our goal is a working product with a compelling end-to-end demo, not disconnected AI features.
PocketQA makes the smartphone part of the intelligence and testing infrastructure.