Blog
Automating Patching for 600+ EC2 Instances
A client ran 600+ EC2 instances across 12 accounts, a mix of Linux and Windows. Patching took 3 senior engineers 4 hours every Sunday. We rebuilt it on AWS Systems Manager and cut the window to 45 minutes.
Killing Terraform Enterprise
We replaced Terraform Enterprise with a free, serverless CI/CD pipeline built on GitHub Actions, S3, and DynamoDB.
The $100k WSO2 Migration
A client's WSO2 API platform was running on 81 EC2 instances with a self-managed HashiCorp stack. We replaced the whole thing with AWS API Gateway. The bill went from about $115k a year to about $3.6k.
Migrating Drupal to WordPress on AWS Lightsail
A client's corporate Drupal site was stable but rigid - every banner change needed a developer. We migrated it to WordPress on AWS Lightsail at $3.50 a month, and their marketing team took the site back.
How a Local Agent Found a 50-Minute Lock, Not a Slow Database
A client called in with 500 errors. The app team said the database was slow. Our local stack, a Hermes Agent running qwen3.8:27b, found a transaction that never committed - in under an hour - and sent two emails. That was the whole bill.
In-House AI for the Support Team
A privacy-focused enterprise had two idle GPUs and a support desk. We turned them into a self-improving, self-hosted support agent - and stopped customer logs from ever touching a third party.
Zero-Cost Domain Redirects with AWS
At sevenseven.tech, we hate waste. We hate wasted CPU cycles, wasted money, and most of all, wasted maintenance time.
Enterprise-Grade Scaling with ECS Fargate & Spot Instances
Enterprise-Grade Scaling with ECS Fargate & Spot Instances. A fully serverless, auto-scaling architecture that balances high performance with aggressive cost optimization.
The Indestructible Frontend Stack
Serverless Hosting with AWS - For the engineer who craves total control, infinite scalability, and near-zero costs, there is no better architecture than the AWS Serverless hexagram: S3, CloudFront, Route 53, ACM, Lambda, and SES.
Optimizing Self-Hosted LLM Inference
Running qwen3.8:27b locally is easy. Making it behave is not. The model card ships explicit sampling profiles for thinking and chat, and the agent profile looks nothing like the low-temperature, high-filter default most stacks ship.
Running a Native AI Stack on Your Own Metal
We swapped Docker Compose for Kubernetes and ran a full AI stack - Ollama, OpenWebUI, Stable Diffusion - on our own workstation. 1,146 tokens/s of prompt processing.
Resurrecting an Old Mining Rig for Night-Shift AI Work
A 2017 crypto-mining rig with four GTX 1060s now runs a full private AI stack. Ollama, OpenWebUI, Docker, and 14 tokens per second on "obsolete" hardware - and why it is perfect for overnight batch jobs.