Skip to content
View Sofille65's full-sized avatar

Block or report Sofille65

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Sofille65/README.md

Sophie — The Monocle Bear

AI Workflows & Agentic UX consultant. Building local LLM infrastructure on Apple Silicon clusters.

Currently running distributed inference on M3 Ultra machines (512 GB + 256 GB × 3) with Inferencer Pro — pushing 570 GB+ models locally, GDPR-compliant, zero cloud dependency.

What I do:

  • Local LLM deployment & multi-tier routing (MLX, Inferencer, Ollama)
  • Agentic UX — systems where AI processes first, humans decide after
  • n8n automation pipelines (audio → Whisper → Claude → Notion)
  • Quantization benchmarking with real-world professional tasks

Stack: MLX · Inferencer Pro · Ollama · n8n · Docker · Cloudflare Tunnels · Tailscale

Previously: 15 years UI/UX, European Commission.

📍 Brussels, Belgium · themonoclebear.com

Pinned Loading

  1. exoscopy exoscopy Public

    Web dashboard for exo, the open-source distributed Apple Silicon inference framework

    HTML 5 2