BRAIN OS INSTITUTE Engineering Blog

Notes from building BOSI and IPJP — brain-cognition assessments that classify 8,192 types. Engineering, research, and product.

Latest

Earlier posts

AI Pacing Debate Heats Up, iOS 27 Launches, and More Tech Headlines
Tech News

AI Pacing Debate Heats Up, iOS 27 Launches, and More Tech Headlines

Dario Amodei calls for slowing AI development, Microsoft publishes MAI rules, SoftBank secures $11.87B for OpenAI, South Korea chips shatter export records, and Marvel Wolverine drops on PS5.

  • #AI Safety
  • #iOS 27
  • #Semiconductors
  • #Startups
GPT-Live API, Asimov Inference Chip, and Key Tech News This Week
Tech News

GPT-Live API, Asimov Inference Chip, and Key Tech News This Week

OpenAI voice AI API, Salesforce job-ready agents, Positron AI challenging NVIDIA inference, and more highlights from this week in tech.

  • #AI
  • #OpenAI
  • #Salesforce
  • #Semiconductor
AI Agent Security Alerts, TSMC Revenue Record, and Key Tech News
Tech News

AI Agent Security Alerts, TSMC Revenue Record, and Key Tech News

Sandbox escape vulnerabilities found in 7 AI coding agents, AI-powered mass PaperCut hack, Meta acquires Stilla.ai, and TSMC posts all-time high monthly revenue.

  • #AI Security
  • #TSMC
  • #Startup Funding
  • #Open Source
NVIDIA Acquires Hugging Face, OpenAI Agents API Launch, and More
Tech News

NVIDIA Acquires Hugging Face, OpenAI Agents API Launch, and More

NVIDIA closes a $12.9B Hugging Face deal, OpenAI ships the Agents API beta, Oracle cloud grows 121% — the tech stories worth your attention this week.

  • #AI
  • #NVIDIA
  • #OpenAI
  • #Cloud
Google's 13B Finland Nuclear Deal, Harvey AI $550M Round, and More
Tech News

Google's 13B Finland Nuclear Deal, Harvey AI $550M Round, and More

Google locks in a 22-year nuclear power deal for Finnish AI data centers, Harvey AI nearly doubles its valuation, NSA warns of Chinese model distillation, and Nintendo Direct drops major reveals.

  • #AI Investment
  • #Google
  • #Nintendo
  • #Open Source
Meta Muse Launch, OpenAI Solves Millennium Prize Problem, and More
Tech News

Meta Muse Launch, OpenAI Solves Millennium Prize Problem, and More

Meta launches Muse personal AI agent, OpenAI claims Navier-Stokes solution, Cognition AI hits $48B valuation, Qualcomm-AWS seal AI chip deal, and other key tech stories for September 10.

  • #Meta Muse
  • #OpenAI
  • #Cognition AI
  • #DeepSeek
Closing the table — stock vs Uncensored Qwen3.8 on Terminal-Bench 89, pass@5 64 to 59
시냅스 일러스트
Engineering

Closing the table — stock vs Uncensored Qwen3.8 on Terminal-Bench 89, pass@5 64 to 59

Round-5 (pass@5) results for the abliterated Uncensored Qwen3.8 Flash Next on Terminal-Bench 2.1, plus a series wrap-up: 59/89 vs stock 64/89, the 13 tasks that split them, speed, and harness failures

  • #terminal-bench
  • #qwen
  • #mlx
  • #abliterated
SoundHound-LivePerson Merger, Foxconn AI Revenue Surge & More
Tech News

SoundHound-LivePerson Merger, Foxconn AI Revenue Surge & More

SoundHound AI closes the LivePerson deal, Foxconn AI servers hit 50% of revenue, CISA flags LiteLLM MCP bypass, Pixxel raises $100M, and Onimusha sells 1M on launch day.

  • #AI
  • #SoundHound
  • #Foxconn
  • #Capcom
NVIDIA Buys Hugging Face, Claude Proves Fermat Theorem
Tech News

NVIDIA Buys Hugging Face, Claude Proves Fermat Theorem

NVIDIA acquires Hugging Face for $12.93B, Claude formalizes Fermat Last Theorem in Lean, and Apple gets a new CEO. Key tech stories this week.

  • #AI
  • #NVIDIA
  • #Hugging Face
  • #Apple
AI Agent Security Funding Surge, NVIDIA NVHBM, and Gaming Big Week
Tech News

AI Agent Security Funding Surge, NVIDIA NVHBM, and Gaming Big Week

From HiddenLayer $100M and Gimlet Labs $300M AI raises to NVIDIA custom HBM, PlayStation State of Play, and Nintendo Direct — this week in tech.

  • #AI Security
  • #NVIDIA
  • #Gaming
  • #Open Source
Crusoe Hits $30B, Tesla Cybercab Under NHTSA Probe, OpenShot 4.0
Tech News

Crusoe Hits $30B, Tesla Cybercab Under NHTSA Probe, OpenShot 4.0

Crusoe triples valuation to $30B, Tesla Cybercab draws an NHTSA probe on launch day, Microsoft ships MAI-Transcribe-2 at $0.10/hr, and more from this week in tech.

  • #Data Centers
  • #Autonomous Driving
  • #Speech Recognition
  • #Open Source
Fable 5.1 Launch, Meta $18B Settlement, AWS DuckLabs Deal
Tech News

Fable 5.1 Launch, Meta $18B Settlement, AWS DuckLabs Deal

Anthropic ships Fable 5.1 with 75% cheaper cache reads, Meta pays up to $18B in child safety settlement, and AWS acquires DuckLabs while DuckDB stays MIT-licensed.

  • #Anthropic
  • #Meta
  • #DuckDB
  • #Firefox
Uncensored Qwen3.8 Flash Next: all 89 Terminal-Bench tasks
시냅스 일러스트
Engineering

Uncensored Qwen3.8 Flash Next: all 89 Terminal-Bench tasks

Three days of Terminal-Bench 2.1 on a community abliterated derivative. Four rounds of context, timeout and temperature changes, 13 server wedges, and what a final 56/89 actually means.

  • #Terminal-Bench
  • #Qwen
  • #MLX
  • #Local LLM
Running Terminal-Bench overnight against Qwen3.8 Flash Next
Summary of an overnight Terminal-Bench run on Qwen3.8 Flash Next — 70/89 completed, 51.4 percent pass rate, timeouts the leading failure cause at 13, and four mlx-serve wedges all recovered by the watchdog
Engineering

Running Terminal-Bench overnight against Qwen3.8 Flash Next

Measuring how much real terminal work a local LLM on an M5 Max can actually complete. The pass rate was less interesting than the server wedging four times and recovering itself every time.

  • #Terminal-Bench
  • #Qwen
  • #MLX
  • #Local LLM
Apple gets a new CEO, DeepSeek Harness hits 112K stars, and other tech news
Tech News

Apple gets a new CEO, DeepSeek Harness hits 112K stars, and other tech news

John Ternus takes over as Apple CEO ahead of the September event, DeepSeek releases its Harness agent framework, six CVEs land in ash_ai, plus OpenMAIC and the ArcBox isolation runtime.

  • #Apple
  • #DeepSeek Harness
  • #AI Security
  • #Open Source
Qwen3.8 27B — 4bit vs 8bit vs Flash Next, and is DFlash2 really lossless?
A comparison of Qwen3.8 Flash Next against 27B at 4bit and 8bit on the same prompts — average decode of 82.5, 65.5, and 18.1 tok/s respectively
Engineering

Qwen3.8 27B — 4bit vs 8bit vs Flash Next, and is DFlash2 really lossless?

Measuring Qwen3.8 27B at 4bit and 8bit against Flash Next, then adding DFlash2 speculative decoding. It was 2–3x faster, but contrary to the official claim the output diverged from the original on M5.

  • #MLX
  • #Qwen
  • #Apple Silicon
  • #Local LLM
Running Qwen3.8 Flash Next on an M5 Max, and the fight with prefill
A before-and-after comparison of tuning Qwen3.8 Flash Next on an M5 Max — cache reuse from 29% to 99.7%, prefill throughput from 354 to 962 tok/s
Engineering

Running Qwen3.8 Flash Next on an M5 Max, and the fight with prefill

Running a 125B MoE model on one Mac as an agent backend. I assumed the model was slow. A day of digging showed the real culprit was prefill — and how caching and a few flags fixed it.

  • #MLX
  • #Qwen
  • #Apple Silicon
  • #Local LLM
Apple ships M6, WRTN becomes a unicorn, Gatik raises $200M, and more
Tech News

Apple ships M6, WRTN becomes a unicorn, Gatik raises $200M, and more

Apples 2nm M6 chip, Intels 256-core Diamond Rapids, WRTN becoming Koreas first AI service unicorn, and Emerald AIs data center power optimization — nine stories from the last week of August.

  • #Apple M6
  • #Intel Diamond Rapids
  • #WRTN
  • #AI Agents
Anthropic wins against the Pentagon, Instinct AI surges, and other tech news
Tech News

Anthropic wins against the Pentagon, Instinct AI surges, and other tech news

A court rules Anthropics blacklisting unconstitutional, the A2A standard consolidates, Instinct AI raises $250M, and sched_ext lands complete in Linux 7.3 — eight stories from the week.

  • #Anthropic
  • #AI Agents
  • #Linux
  • #Instinct AI
Buzz — the open-source workspace that gives agents an identity
Tech News

Buzz — the open-source workspace that gives agents an identity

Block released Buzz, which treats AI agents as workspace members rather than add-ons. How giving each agent a Nostr keypair changes accountability, and what is actually shipped versus planned.

  • #Open Source
  • #AI Agents
  • #Nostr
  • #Block
NVIDIA posts $96.2B, OpenAI unveils its own chip, and other tech news
Tech News

NVIDIA posts $96.2B, OpenAI unveils its own chip, and other tech news

NVIDIA reported record quarterly revenue and moved on Hugging Face while OpenAI claimed its in-house inference chip beats Blackwell. A roundup of the weeks major moves across the industry.

  • #NVIDIA
  • #OpenAI
  • #Meta
  • #Amazon
[postgresql] pg_wal that will not shrink — start with replication slots
Bar lengths contrasting a 650MB database against an 18GB pg_wal directory
Engineering

[postgresql] pg_wal that will not shrink — start with replication slots

PostgreSQL will not delete WAL while a replication slot still holds it. Two months of WAL accumulated after a standby died, and the secondary damage that only surfaced once the space came back.

  • #PostgreSQL
  • #Operations
  • #Monitoring
Career Counseling — BOSI for adults and teams
A team in a meeting
Solutions

Career Counseling — BOSI for adults and teams

What changes when the same brain-cognition assessment is applied to working adults and companies instead of students. Career design, hiring, team composition, and finding a co-founder.

  • #BOSI
  • #Career Counseling
  • #Hiring
  • #Teams
Edu Counseling — BOSI, rebuilt for parents
An illustration of a brain-cognition network
Solutions

Edu Counseling — BOSI, rebuilt for parents

Our BOSI brain-cognition assessment, packaged so that parents of school-age children can act on it. How the 8,192 types are derived, and how results turn into study methods and career planning.

  • #BOSI
  • #Edu Counseling
  • #Solutions
Starting an engineering blog
회의 중인 기업 팀
Updates

Starting an engineering blog

We build brain-cognition assessments that sort people into 8,192 types. This is where we write down the problems we hit along the way and the choices we made.

  • #Announcement
[proxmox] Fixing "CPU does not support x86-64-v2"
The default CPU model kvm64 failing to meet the x86-64-v2 level a container image requires
Engineering

[proxmox] Fixing "CPU does not support x86-64-v2"

Modern container images require x86-64-v2. The default CPU model your hypervisor hands a VM does not have those instructions, so the process dies inside glibc before the application ever runs.

  • #Proxmox
  • #Virtualization
  • #Containers
Fixing ssh "Connection timed out during banner exchange"
A contrast showing the SYN packet of a connection passing while the ACK packet of the same connection is dropped
Engineering

Fixing ssh "Connection timed out during banner exchange"

If your iptables whitelist only inspects ctstate NEW, the handshake never completes. A port scan still reports the port open, which is why the cause of a failing backup went unfound for a month.

  • #iptables
  • #Networking
  • #Troubleshooting
[haproxy] The reload succeeded and the config still did not change
A reload command returning rc=0 and OK while the configuration remains unchanged
Engineering

[haproxy] The reload succeeded and the config still did not change

A reload can return rc=0 without applying anything. Two patterns: an old process surviving and serving the previous config, and changes landing in a staging file that never got promoted.

  • #HAProxy
  • #Troubleshooting
  • #Operations
Recovering a failed Ubuntu 22.04 → 26.04 upgrade without reinstalling
A decision cue showing that half-installed at zero means no reinstall is needed, and 750 unpacked packages dropping to zero
Engineering

Recovering a failed Ubuntu 22.04 → 26.04 upgrade without reinstalling

Skipping an LTS step and stalling at the configure stage takes sudo and networking with it. Counting dpkg states tells you whether recovery without a reinstall is possible.

  • #Ubuntu
  • #Linux
  • #Recovery
[monitoring] When "last updated at" is the wrong thing to alert on
A side-by-side contrast between a metric that stays still when healthy and one that moves when work is happening
Engineering

[monitoring] When "last updated at" is the wrong thing to alert on

If you alert on a value that does not change while things are working, it will fire eventually — guaranteed. The metric that produced false alarms, and the real alert we silenced while fixing it.

  • #Monitoring
  • #Operations
  • #Alerting
[linux] You changed the timezone, but the DB and containers did not
Three boxes contrasting an OS set to KST against a database and container still running UTC
Engineering

[linux] You changed the timezone, but the DB and containers did not

timedatectl changes the OS timezone. Your database and your containers keep the old one. How NOW() came back nine hours apart inside one cluster, and why nothing short of a restart fixes it.

  • #Linux
  • #MariaDB
  • #Docker