Your Own Private AI on a $798 Laptop: 2B Model, 91% Pass
We built an AI router that keeps 89% of queries free and on-device. 35 tests. Real numbers.
Educational demos, study builds, and technical experiments by NosisTech. Published for learning and documentation only, not as production systems or implementation guidance.
We built an AI router that keeps 89% of queries free and on-device. 35 tests. Real numbers.
A $798 gaming laptop with a 4GB GPU runs private local AI and automatically escalates the hard stuff to the cloud. Here’s the test data.
We installed a local AI on a $798 gaming laptop and ran 71 real-world tasks.
You changed the context to 64k and Hermes still rejects it. Here is the parallel slots problem nobody is writing about yet.
How I turned a 12GB laptop GPU into a 114 tok/s local AI agent with MTP, Hermes, and three config changes.
An educational demo inspired by OpenHands, showing review-gated software task planning with LiteLLM.
An educational demo of an agent that checks AI responses against source context and flags unsupported or contradicted claims.
An educational demo of a LiteLLM-native evaluation harness that tests AI responses with visible assertions and compact verdicts.
A safe educational rebuild of the LLM Guard pattern as a small LiteLLM safety gateway that scans inputs, redacts data, and checks outputs.
educaitonal demo of small prompt injection screening agent that flags suspicious text before it reaches a main AI model.