GitHub – stanford-mast/blast: Browser-LLM Auto-Scaling Technology

by oqtey May 2, 2025

written by oqtey May 2, 2025

GitHub - stanford-mast/blast: Browser-LLM Auto-Scaling Technology

A high-performance serving engine for web browsing AI.

I want to add web browsing AI to my app… BLAST serves web browsing AI with an OpenAI-compatible API and concurrency and streaming baked in.
I need to automate workflows… BLAST will automatically cache and parallelize to keep costs down and enable interactive-level latencies.
Just want to use this locally… BLAST makes sure you stay under budget and not hog your computer’s memory.

pip install blastai && blastai serve

from openai import OpenAI

client = OpenAI(
    api_key="not-needed",
    base_url="http://127.0.0.1:8000"
)

# Stream real-time browser actions
stream = client.responses.create(
    model="not-needed",
    input="Compare fried chicken reviews for top 10 fast food restaurants",
    stream=True
)

for event in stream:
    if event.type == "response.output_text.delta":
        print(event.delta if " " in event.delta else "", end="", flush=True)

🔄 OpenAI-Compatible API Drop-in replacement for OpenAI’s API
🚄 High Performance Automatic parallelism and prefix caching
📡 Streaming Stream browser-augmented LLM output to users
📊 Concurrency Out-of-the-box support many users with efficient resource management

Visit documentation to learn more.

Awesome! See our Contributing Guide for details.

As it should be!

3D printing 3D scanning 5G 6G Adaptive learning AI AI ethics AI governance AI-driven automation AI-driven chatbots AI-driven healthcare AR/VR (Augmented and Virtual Reality)Artificial intelligence Augmented reality Automation Autonomous drones Autonomous vehicles Big data Bioinformatics Biometric security Blockchain Blockchain security Blockchain-as-a-Service Chatbots Cloud computing Cloud infrastructure Cloud security Cloud-native applications Cognitive computing Cryptocurrency Cyber defense Cyber-physical systems Cybersecurity Cybersecurity frameworks Data analytics Data governance Data lakes Data mining Data privacy Deep learning DevOps Digital currency Digital ecosystems Digital payments Digital transformation Digital twins Digital wallets Drones Edge AI Edge computing eSIM technology Fintech Fintech innovation Geospatial analytics Gig economy platforms Green technology Human augmentation Hybrid cloud Hyperautomation Image recognition Intelligent apps Internet of Behaviors (IoB)IoT (Internet of Things)IT operations IT security Machine learning Metaverse Microservices Mobile app development Multi-cloud environments Multi-factor authentication Natural language processing Neural networks Open-source software Predictive analytics Privacy-enhancing technologies Quantum computing Quantum encryption Quantum sensors Renewable energy storage Renewable energy tech Robotics Robotics process automation (RPA)SaaS (Software as a Service)Self-driving cars Serverless computing Smart cities Smart contracts Smart devices Smart grids Smart homes Supply chain tech Tech sustainability Video streaming Virtual assistants Virtual reality Voice recognition Wearable health tech Wearable technology Zero-trust security

GitHub – stanford-mast/blast: Browser-LLM Auto-Scaling Technology

The Final Reckoning’ & More

Wall Street and European markets finish week on a high after US jobs report | Stock markets

Related Posts

Leave a Comment Cancel Reply