bbioonThemes
  • Home
  • Blog

Tag: AI Assistants

AI, AI in WordPress, Development

NeMo Agent Toolkit: real metrics for LLM agents

Most LLM agents ship with no visibility between the input and the answer. How the NeMo Agent Toolkit traces every tool call and scores trajectory rather than answer accuracy alone, using Arize Phoenix and W&B Weave.

Read Article
AI, Development

How I use slash commands in Cursor and Claude Code

I stopped retyping the same AI prompts and moved them into a commands folder. How slash commands work in Cursor and Claude Code, the two I use most on WordPress projects, and why that folder belongs in Git.

Read Article
AI

Automatic prompt optimization for vision agents with HRPO

Manual prompt tweaking does not survive contact with vision models. A walkthrough of HRPO in the Opik-optimizer SDK on driving data: how the loop is wired, why you need a hold-out set, and the run that took accuracy from 15% to 39%.

Read Article
AI, Design, Development

Penpot MCP server: design context an AI can read

Penpot is testing an MCP server that lets assistants like Claude read real design files instead of guessing from your prompt. What is in the experiment, how the translation layer works, and why design-as-code cuts down on invented components.

Read Article
AI, AI in WordPress, Development

AI coding agent context: give it a map, not your whole repo

Three habits that stop a coding agent inventing hooks and tables: an AGENTS.md file it has to keep updated, real documentation and schema files in the context, and a fresh thread when the task changes.

Read Article
AI, AI in WordPress

Where AI in UX actually saves me time

Two years of daily use, boiled down: AI in UX is good at sorting research transcripts, flagging the obvious usability problems and drafting copy from client bullets. The judgment stays with you.

Read Article
AI, Development

RAG chunk size is the variable most people never tune

Small chunks lose context, medium chunks produce near identical similarity scores, large chunks are stable but noisy. How RAG chunk size affects retrieval, plus a PHP splitter that respects word boundaries.

Read Article
AI, Development

Reinforcement learning for LLM reasoning: proofs over vibes

Part 2 of the vibe proving series: how an RL loop built around a hard proof checker teaches a model to follow logic rules, and where it still falls over on nested proofs.

Read Article
AI, Development

Why I use GliNER2 for structured data extraction

A client had ten thousand messy biographical records that needed to become a clean knowledge graph. GPT-4o would have done it and cost a fortune, so I used GliNER2 instead: schema-driven extraction that runs on a CPU.

Read Article
AI, Development

How I handle covariance shift with IPW

A WooCommerce recommendation engine scored 0.85 AUC in staging and fell apart on launch day because the live audience did not match the training data. Inverse probability weighting is how I got an honest read on what the model was really doing.

Read Article

Posts navigation

Previous 1 … 15 16 17 … 22 Next

bbioonThemes

Senior WordPress Engineer
Toptal & Codeable Expert

Connect

  • GitHub
  • Twitter
  • LinkedIn
  • Codeable

Explore

  • Expertise
  • Work
  • Insights
  • Blog

Resources

  • bbioonThemes
  • Contact
  • Privacy Policy
© 2026 bbioonThemes. All rights reserved.
Privacy