AI/Automation Engineer

Building
production apps
through vibe coding

Turning business problems into production apps through AI-assisted development. Full cycle: from idea to production.

30+
projects built
24
in production
75+
Docker containers
15+
LLM models in stack
~/projects — claude code
live
PAUSED
vibe-coding session
cosmodoc 
scroll
Claude Code React FastAPI Docker Next.js PostgreSQL Dify NestJS Playwright Prisma Flask TypeScript Claude Code React FastAPI Docker Next.js PostgreSQL Dify NestJS Playwright Prisma Flask TypeScript

Vibe Coding —
not just hype

I describe functionality in natural language, AI generates code, then iterative debugging and deployment. This approach allows building fullstack apps with microservice architecture in days, not months.

8+ years in marketing and business give me deep understanding of client needs. Enterprise projects for major international brands, B2B platforms, 75+ Docker containers across a NAS and two VPS. I build not just technical solutions, but tools that solve real business problems.

OPEN_FOR_WORK
AI ENGINEERING
& AUTOMATION

Full-cycle development: from business analysis to production deployment. AI integrations, microservices, self-hosted infrastructure.

AI DEVELOPMENT VIBE CODING AUTOMATION API INTEGRATION CONSULTING
Book an Interview ↗
THROUGHPUT
Technologies

Tech stack
powering production

AI / LLM
Claude Sonnet 4.6Claude Haiku 4.5GPT-4o Visiongpt-image-1Gemini 2.5 FlashOpenRouterDify 1.11RAGWeaviatepgvectorfastembedOllamaMCPYandexGPTGigaChat-2KandinskySuno AIfal.ai FLUXWhisper / PiperPrompt Engineering
Vibe Coding
Claude Code CLIClaude Agent SDKMulti-Agent SystemsMCP serversPRP FrameworkBMAD MethodAuto-Claude
Backend
Python 3.12FastAPIFlaskNestJS 11Express.jsNext.js 16Node.jsTypeScriptLaravel 11StreamlitPydanticSQLAlchemyaiogram 3tRPC v11
Frontend
React 19TypeScriptTailwind CSS v4shadcn/uiVite 7React NativeExpoHTMXRechartsPlotlyElectronPWA
Databases
PostgreSQL 18DuckDB + ParquetSupabaseMySQLMSSQLSQLiteRedisWeaviateMinIOPrisma 7Alembic
DevOps
DockerDocker ComposenginxLet's EncryptsystemdUbuntuHeadscaleWireGuardAmneziaWGHysteria 2CloudflareUptime Kuma
Automation
n8nPlaywrightAPSchedulerETLCron WorkersCloudflare WorkersFlowise
APIs & Protocols
RESTSOAPMTProtoWebSocketSSESSH TunnelsTelegram Bot API
Portfolio

Production projects

Email Digest Generator

AI email marketing automation
Production

Automated monthly email digest generation platform for a leading electronics manufacturer. Full cycle: content scraping → AI variant generation → analytics scoring from 262 templates → editor UI → designer export. Separate PROD version on Russian AI (YandexGPT, GigaChat) for 152-FZ compliance.

Python 3.12FlaskPostgreSQLClaude Sonnet/Haikugpt-image-1PlaywrightPillowYandexGPTGigaChat
AI generation: 5 headlines + summaries + clickers + AI images
Analytics Ranker (0-100 scoring from 262 templates, 210+ campaigns)
Step-by-step wizard: 40+ routes, inline editing
Outlook-compatible template (table layout, MSO conditionals)
Russian AI version for 152-FZ (YandexGPT, GigaChat, Kandinsky)
DB: 12 tables, 49 pytest tests, AI cost tracking

AI Agents Platform

Multi-agent system + custom microservices
Production

Multi-agent platform on Dify with two specialized AI agents for an electronics manufacturer's CRM team. Custom microservices: File Storage API (16 endpoints) + Analytics API (12 endpoints). Anti-hallucination design, cross-chat memory, 24 tools. Prompt optimization cut conversation cost by 85% ($1.40 → $0.22). A Russia-only deployment runs entirely on GigaChat-2 + YandexGPT — no foreign APIs.

Dify 1.11Claude Sonnet 4Flask ×2PostgreSQLRedisWeaviateGPT-4o-miniDocker
"Project Evaluator" agent: 3 modes, 34 KB prompt
"Universal Analyst" agent: 5 analysis modes
Anti-hallucination: 4 source levels, PII validation (7 types)
File Storage API: logs, briefs, employees, knowledge, summaries
Analytics API: 12 endpoints, 253 campaigns, 465+ articles
Prompt optimization: 5143 → 1449 tokens (−72%), cost −85%
Russia-only deployment: GigaChat-2 + YandexGPT Lite (100% RU models)
40+ Docker containers in the stack

Email Analytics

AI email campaign analytics
Production

AI-powered email campaign analytics platform for a leading electronics manufacturer. Analysis of 620+ campaigns, ~6000 click elements, AI insight generation, branded PPTX/PDF reports and click-map visualizations.

Python 3.12StreamlitSupabaseClaude APIGPT-4 VisionpandasPlotlypython-pptxReportLab
620+ campaigns, ~6000 click elements, 120+ HTML designs
Vision AI for email design analysis (6 criteria)
Auto-generated PPTX/PDF with corporate branding
AI chat: questions about campaigns in natural language
Click-map visualizations and heatmaps
ETL pipeline (Excel + ZIP → AI → reports)
View demo
Email Analytics Dashboard
Email Analytics Campaigns
Email Analytics Analytics
Email Analytics AI Chat

CDP LeadGen

B2B lead generation platform
Production

Microservice platform for automated lead discovery and AI scoring for CDP sales. Aggregates signals from 4 sources, applies multi-level scoring with time decay.

Next.js 16React 19TypeScriptPythonTelethonSupabaseClaude SonnetDocker Compose
Microservice architecture (Dashboard + Telegram Parser + Cron Worker)
4 sources: HH.ru API, Kontur.Zakupki SOAP, RSS, Telegram MTProto
AI Bulk Analysis: mass analysis via Claude with SSE progress
Multi-product architecture with data isolation
Rating system: QuickVote, StarRating, AI vs Human correlation
Tender document parsing (PDF/DOCX)
View demo
CDP Dashboard
CDP Companies
CDP Analytics

HR Automate

Recruitment automation platform
Production

Full-stack HR automation platform for a full-service marketing agency. 11 modules, 9 cron jobs, 5 OAuth integrations. Parallel EU and Russian environments routing between Western and Russian AI models based on PII presence (152-FZ compliance).

FastAPINext.js 15PostgreSQLClaude Sonnet 4YandexGPTHH.ru OAuthMS GraphTelethonDocker
11 modules: market analytics, AI messages, booking, resume eval, PII
5 OAuth integrations: HH.ru, Huntflow, MS Graph, Telegram MTProto, O365
Security-driven AI routing: PII → YandexGPT, anonymized → Claude
AI generator for HH boolean queries (noise reduction from ~6 000 to relevant subset)
9 cron jobs (MSK), atomic UPDATE...RETURNING for de-duplication
XLSX market report: 4 regions × Min/Mid/Max × 3 experience tiers

CRM Bulk Analytics

Multi-channel marketing analytics
Production

Unified analytics platform for bulk communications (email + Telegram + MAX) for a leading international FMCG holding — multiple brands in one system. Replaced 3 manual PPTX reports. Brief Oracle predicts OR/CTOR/Unsub for new campaigns via similarity search through pgvector HNSW. Version v11.5: 334+ campaigns; iterative calibration of AI insights through analyst feedback cut critical-rule violations from 73% to 4%.

FastAPINext.js 15PostgreSQLpgvector HNSWClaude Sonnet 4.6Rechartspython-pptxDocker
5 sections: Dashboard (8 KPIs), Campaigns, Compare, Brief Oracle, Reports
Comparability scoring 0-100 (apples-to-apples vs vinegar warning)
Brief Oracle: top-20 similarity → CI95 ErrorBars → ship/iterate/hold
AI prompt v8: 70/70 coverage, 127 PPTX few-shot examples, RAG from 25 chunks
v11.5: 334+ campaigns, critical-rule violations 73% → 4% (5 calibration rounds)
Vision pipeline: 156/156 TG campaigns with caption extraction from screenshots
9.7M rows of CRM communications in a linked DuckDB/Parquet lake
Branded PPTX: python-pptx XML patches, adaptive font sizing

SMS Analytics

Messaging analytics platform
Production

Automated collection, classification and reporting of SMS/Telegram messages across multiple brands. Replaced 8 n8n workflows with a single FastAPI service. 6 SMSC accounts, 7 brands, ~10,000 messages/week.

Python 3.11FastAPISQLAlchemyPostgreSQL 16HTMXAPSchedulerAlembicDocker
Auto-classification: 5 types (welcome, mgm, dormant, service, bulk)
Scheduler: collection (Mon 07:00) → report (Mon 10:00) → email
Password-protected ZIP reports with 24h download links
HTMX interface with no page reloads

Analytics MCP

MCP server for an analytics team (DuckDB lake)
Production

MCP server connecting arbitrary CSV/XLSX/Parquet datasets up to 1+ GB to Claude Teams as a Custom Integration. 9.7M rows of CRM communications in a DuckDB/Parquet lake plus Slide RAG over 649 corporate presentation slides. Analysts query data in natural language.

FastMCPDuckDBParquetOAuth 2.1 + DCRfastembedClaude TeamssqlglotDocker
9.7M rows: CSV 5.6 GB → Parquet 68 MB (84× compression) in a 4-min ETL
13 MCP tools: data lake (10) + Slide RAG (3, hybrid RRF: vector + FTS)
OAuth 2.1 + Dynamic Client Registration (RFC 7591/8414/9728)
multilingual-e5-large embeddings + DuckDB FTS with Russian stemmer
Vision captioning of 649 slides (Claude Sonnet 4.6, offline indexing)
SQL validation via sqlglot (Select/Union/With only — injection guard)

Email Program Analytics

Editorial email-program analytics (FMCG, multi-brand)
Production

Editorial-grade analytics platform for the email program of a leading international FMCG holding (7 brands): 1,433 templates, 27 months of data, 19.27M sends. Next.js + FastAPI fully replaced the original Streamlit prototype.

Next.js 16React 19FastAPIPostgreSQLPlaywrightshadcn/uiClaude Sonnet 4.6Tailwind 4
Brief Oracle: OR/CTOR prediction + SCoP (Pearson r=0.347)
Heatmap render via headless Playwright over 1,198 email screenshots
Tag effectiveness: "money" +13.2pp CTOR, personalization −4.5pp (robust)
Auto-tagger: 54 tokens → 5 categories with an alias system
PPTX Smart Comparison with match-score (brand + segment + date + volume)
AI hand-tests: GigaChat-Max ≈ Claude Sonnet (8/10)

Report Automate

Agency management reporting
Production

Internal management-reporting system for a marketing agency: estimates, hours close-out, margins, OPEX, MReport — replacing manual Excel. 23 entities, 37 screens, 5 reports plus an AI reporting consultant.

FastAPINext.js 16React 19PostgreSQLpsycopg2shadcn/uiClaude Sonnet 4.5systemdDocker
100% spec coverage: 23/23 entities, 37/37 screens, 5/5 reports
Raw psycopg2 + RealDictCursor for complex financial queries
RLS: a group lead sees only their own clients
systemd-timer for monthly recalculation
AI reporting consultant (Claude Sonnet 4.5 / Haiku 4.5)

PBI Anomaly Detector

Power BI anomaly detector
Production

Automated anomaly detector for Power BI metrics at an electronics manufacturer: monitors sales channels and customer segments in an MSSQL source and alerts in Aspro Cloud with an @mention of the responsible manager.

PythonMSSQL / pyodbcAspro Cloud APIWindows Task SchedulerVPN
Two strategies: presence check + deviation >20% from 15-day average
MSSQL read-only over VPN, daily 10:00 MSK
Aspro Cloud: auto-tasks with @mention, idempotency via persistent state

Onboarding Generator

Onboarding email-series generator
MVP

Generator of onboarding email series for buyers of flagship smartphones (day 7/14/30). Playwright scrapes the manufacturer's catalog, AI drafts 2–3 emails, export as a ZIP of HTML.

FlaskPostgreSQLPlaywrightClaude Sonnet 4Claude HaikuOpenRouterJinja2
Playwright catalog scraping: products, specs, highlights
Claude Sonnet 4 (subjects) + Haiku (sections) via OpenRouter
600px Outlook-compatible templates via an assembler
Editorial flow: pick variants, preview, ZIP export

Figma Brief Automation

Auto-generated GA briefs from Figma
MVP

CLI tool and MCP server that auto-fill an Analytics Brief (GA events) from Figma designs. Analysts used to spend hours manually describing screens — the tool automates the task.

Python 3.12Figma APIMCPopenpyxlClaude API
Analyzes Figma frames via API, classifies screens and buttons
Fills the Analytics Brief XLSX template + generates GA events
MCP server for integration with Claude Code

Investment Research Digest

Investment digest generator
In progress

Generator of email digests from investment research for qualified investors of a Russian investment firm. Playwright scrapes a Bitrix SPA and re-renders Highcharts charts via plotly in a branded style.

FlaskPostgreSQLPlaywrightplotly + kaleidoClaude Sonnet 4Claude Haiku 4.5OpenRouter
Playwright parses the Bitrix SPA, extracts Highcharts graphs
plotly + kaleido: re-render to PNG in a branded b/w premium style
Interactive chart picker: original Highcharts or plotly re-render
Isolated fork of the main Digest Generator

AI Sales System

Lead architect · manufacturing company
In progress

A self-hosted pipeline for inbound sales requests: classified-ad messages, website forms, and calls all pass AI qualification and automatic shipping-cost calculation, then land in the CRM as ready leads with a drafted reply to the buyer. My role is lead architect — CRM audit, access hand-over, architecture, code, and the admin panel.

PythonFastAPIPostgreSQLYandexGPTOllamaBitrix24 APIDocker Composepytest
Work queue inside PostgreSQL itself (FOR UPDATE SKIP LOCKED) — no Redis
Client compliance: self-hosted, Russian and local LLMs only
Shipping cost quoted live from three carrier APIs
"Client zone": prompts, reply templates, and regions editable without a developer
500+ tests against a real PostgreSQL, migrations under an advisory lock
Separate dev environment with public HTTPS for webhook intake

Committee AI Moderator

Tenant matching for a commercial real estate agency
Pilot

A Telegram bot that runs the agency's weekly tenant-matching "committee" instead of manual reconciliation. It pulls new properties from the CRM, an LLM matches them against open applications, brokers reply right in the chat, the bot parses those replies, checks the promises against the CRM, and posts a ✅/❌ report the next business day.

FastAPIaiogram 3SQLAlchemyAPSchedulerClaude HaikuGeminiYandex STTWhisperDocker
Deterministic prefilter before the LLM (use, area ±20%, budget, district) keeps token cost bounded
Broker replies: the LLM extracts, code normalizes (rapidfuzz) — not the model
Voice feedback from the chair, transcribed by three STT providers side by side
Few-shot "lessons from past committees" built from accumulated feedback instead of fine-tuning
Evening re-check plus a penalty draft behind an admin-only button
Mock CRM for demos and the live CRM behind one interface, tests 24/24

Buildify

Home services mobile marketplace (UAE)
Production

A home services marketplace for the UAE market with three roles — customer, executor, mediator. I worked on production hardening and shipping the app to Google Play and the App Store: React Native upgrade, release builds, payments, monitoring.

React Native 0.76TypeScriptReduxLaravel 11PHP 8.2PostgreSQL 15RedisSoketiStripeFirebaseSentry
React Native 0.75 → 0.76 upgrade, release AAB/IPA builds
Three UI languages (RU/EN/AR) via i18next, including RTL
Real-time chat and notifications over Soketi/Pusher
Stripe Cashier payments, Firebase analytics, Sentry monitoring
AI design-concept generation: text prompts + Stable Diffusion SDXL img2img
Builds shipped to TestFlight and Google Play Internal Testing

Claude Workspaces

Multi-user AI platform
Production

Platform providing isolated Claude Code environments (VS Code in browser + file manager) for 10 concurrent users with a library of 19 skills.

code-serverClaude Code CLIDocker ComposePythonNode.js 22
10 isolated environments (0.5 CPU, 768 MB RAM per user)
19 skills: code, documents, design, translations
3 AI modes: Claude OAuth, Router (GLM-5), Z.AI
Auth-gateway + File Browser per user

CosmoDoc

Medical practice management PWA
Production

Mobile-first Progressive Web App for cosmetology clinic management. Patient record sync from YClients, visit tracking with before/after photos, doctor earnings calculator.

React 18TypeScriptViteExpress.jsPrismaSQLiteYClients APIDocker
Bidirectional YClients sync (30 min intervals)
PIN protection (lockout after 5 attempts)
Earnings calculator (9 service categories)
Voice input via Web Speech API
Patient before/after photo gallery
Excel export for accounting
View demo
CosmoDoc Dashboard
CosmoDoc Patients
CosmoDoc Finances
CosmoDoc Statistics
CosmoDoc Analytics

Plant Care

AI-powered plant care PWA
Production

Progressive Web App with AI recommendations for houseplant care. Watering schedules with push notifications, plant identification, personalized tips from Claude Haiku.

Next.js 16React 19tRPC v11Prisma 7PostgreSQLClaude HaikuDocker
AI care recommendations (Claude Haiku)
PWA with push notifications for watering
Personalized care schedules
Public access: plant.a-van.info

Chronicler's Forge

D&D 5e Digital Tabletop
Development

Web app for running D&D campaigns: digital character sheets, maps with markers, Suno AI music, fal.ai image generation, Telegram content import. Real-time sync between DM and players.

NestJS 11Prisma 7PostgreSQLRedisMinIOReact 19Socket.IOSuno AIfal.ai
D&D 5e character sheets (stats, spells, inventory)
Maps with SVG markers (16 types) and legend
Suno AI music generation (28 D&D presets, 6 models)
fal.ai image generation for lore and spells
Adventure Templates: 5 parsers (JSON, MD, PDF, DOCX, Foundry VTT)
Telegram content import (MTProto, ~4800 messages)

NPS Classifier

OpenAI-compatible API proxy
Production

Multi-backend LLM proxy for NPS comment classification. Supports OpenRouter, Z.ai, LM Studio with automatic fallback and benchmarking.

Python 3.12FastAPISQLiteOpenRouterZ.aiDocker
OpenAI-compatible API (/v1/chat/completions)
Per-request backend override (X-Backend header)
Benchmark endpoint: /v1/compare across all backends
Automatic fallback on failures

hh-auto

Job search automation
Production

Full-cycle job search automation on hh.ru: vacancy search via API, AI scoring (Gemini), cover letter generation (Claude Sonnet), automated application via Playwright.

FastAPIPlaywrightPostgreSQL 16HTMXClaude SonnetGeminiDocker
AI pipeline: fast_score → ai_score (Gemini) → cover letter (Claude)
7 automated scheduled tasks
Best resume selection from 5 variants per vacancy
Closed API workaround via Playwright browser automation

AI Caller

AI-powered calling system
Development

Automated calling system with AI: speech recognition (Yandex STT), response generation (OpenRouter AI), speech synthesis (Yandex TTS). Real-time voice processing via WebSocket.

FastAPIPlaywrightReactYandex STT/TTSOpenRouterWebSocketDocker
Real-time voice processing via WebSocket
Yandex STT (recognition) + TTS (synthesis)
On-the-fly AI response generation

fit-tracker

Training PWA with an AI coach
Production

A home strength-training program turned into an app: set logging, wellbeing tracking, reminders, and an AI coach that adjusts the program in conversation. The core of the project is a machine-enforced safety layer — the user has an injured knee, and the constraints live in code rather than in the text of the program.

Next.js 16tRPC v11Prisma 7PostgreSQL 18Better AuthSerwist PWAweb-pushClaude APIDocker
Safety gate: the program seed fails validation if an exercise violates the knee constraints
The same rules apply to AI-coach edits — the model cannot route around a constraint
AI coach in chat: Claude API with SSE streaming
Bearer-token agent API: an external agent reads and updates the program on a schedule
Push reminders: morning workout, supplements, hourly movement breaks during the workday
Live at fit.a-van.info

Shinugi

Interactive textbook for a drum school
Production

A drum school's printed textbook rebuilt as an interactive app. One codebase produces both the web build and the mobile app. Notation is rendered by an engine from typed content rather than shipped as images — so a mistake in the material is caught by the compiler.

Expo SDK 56React Native 0.85React 19expo-routerreact-native-webTypeScriptJestnginx
One codebase: static web export plus a mobile app
SVG notation engine: note durations, beats, staff layout
Content is typed — a schema error fails typecheck instead of reaching a student
Source pipeline: textbook PDF and TIF artwork → webp + an asset manifest
Live at drums.a-van.info

Ad Balance

Ad account balance monitoring
Production

A dashboard for balances and spend across ad accounts on three platforms. It replaced an n8n workflow with a dedicated app: balance history, per-campaign spend breakdown, and alerts as an account approaches zero.

Next.js 16React 19tRPC v11Prisma 7PostgreSQL 18RechartsTailwind v4Docker
Three integrations: VK Ads, Unity Ads, Mintegral
Polling scheduler running inside the app every 10 minutes
Balance history, daily spend, per-campaign breakdown
Slack alerts; platform credentials live in settings, not in code

Laser Calc

Laser cutting cost calculator
Production

An industry calculator for laser cutting cost: material, thickness, cut and pierce length, mode and speed in — cost and machine time out. Built for a specific shop floor and running publicly.

Next.js 16React 19TypeScriptTailwind CSS v4Dockernginx
Cost model driven by material, thickness, and cutting mode
Live at laser.a-van.info

Family AI Agents

Deploying and extending AnthroClaw (open source)
Production

Three personal Claude agents for family members behind a single Telegram bot, each with its own memory, permissions, and skills. The agents drive the home media server — they take a movie or album request by voice or text, queue it, and report back on their own once it is ready.

Claude Agent SDKNode.js 22Telegram Bot APIMCPDockernginxWyoming
Agent isolation: separate workspace, prompt, memory, and permissions per agent
Requests routed back to their author through media-stack tags
Event-driven notifications instead of polling: the agent writes when the request completes
Voice loop: smart speaker → Whisper → agent → Piper and back
Container self-healing and a gateway watchdog

Private Network & VPN Stack

Self-hosted coordinator, two VPS, censorship resistance
Production

The infrastructure layer underneath every other project: a private overlay network on my own coordinator instead of a cloud one, a resilient egress path, and public exposure of home services without a single port forward on the router.

HeadscaleWireGuardAmneziaWGHysteria 2nginxLet's EncryptSSH tunnelsUptime Kuma
Self-hosted private-network coordinator (Headscale) plus an own relay
Migration off TLS-based protocols to UDP after filtering throttled them to 500 B/s
Two independent VPN protocols: primary and fallback
SSH reverse tunnels publish home services through a VPS
Cascade through a second VPS plus self-service device onboarding
Monitoring: health checks, auto-restarts, Telegram alerts

Wedding Invitation

One-page invitation site with RSVP
Production

A single-page invitation site with its own visual identity: an arched window frame, botanical dividers, and a vertical date rail. The greeting is personalized from the link, and RSVP answers arrive in Telegram.

ViteTypeScriptCSSTelegram Bot APInginx
All content lives in one config file — no markup editing needed
Per-guest personalized links via a query parameter
RSVP form posts answers straight into Telegram
Live at wedding.a-van.info
Experience

Career path

2026 — present
Lead Architect, AI Sales System
Manufacturing company (contract)
Owning the project end to end: CRM audit and access hand-over, architecture of a self-hosted "inbound request → AI qualification → CRM lead" pipeline, development, admin panel, and handover to the client. Compliance constraint: Russian and locally hosted models only.
2024 — present
AI/Automation Engineer & Vibe Coder
Independent practice
Full-cycle production app development through AI-assisted development. 30+ projects, 24 in production, 75+ Docker containers across a NAS and two VPS.
2022 — present
Marketer → No-code/AI Developer
Russian financial holding
Evolution from marketing to development. Enterprise AI platforms (email automation, AI agents, analytics), 100+ n8n automations, Dify Platform with RAG pipelines.
2018 — 2022
Head of Marketing & Advertising
Irish pub chain
Full-cycle marketing for a restaurant chain. Deep understanding of business processes, now applied to designing AI solutions.
2008 — 2018
Marketing, media, content
Various companies
10 years in marketing, media and content — from journalism to advertising budget management. Foundation for understanding business needs.
Approach

Full cycle:
from idea to production

01

Idea Analysis

Deep dive into business context, problem and goal definition. 8+ years in marketing enable instant understanding of client needs.

02

Design & Planning

User scenarios, wireframes, success metrics and MVP scope. Defining what to build first.

03

Architecture & Stack

Technology selection, data schema, API design, infrastructure decisions. Docker, microservices, 30+ API integrations.

04

Vibe Coding & MVP

AI-assisted development: prompt → code → iterative debugging → working prototype in days, not months.

05

Scaling & Optimization

Load testing, caching, query optimization, refactoring. Preparing for user growth.

06

Production & Support

Docker deployment, SSL, monitoring, health checks, alerting. Full responsibility for production service operation.

Automation

100+ automations
& integrations

01

YClients ↔ CRM Bridge

Bidirectional patient and booking sync between YClients and internal CRM every 30 min. 9 service category mapping, deduplication, conflict resolution.

n8nYClients APIWebhookPrisma
02

AI Document Pipeline

ETL: PDF/Excel upload → parsing → Claude API analysis → structured data → auto-generated PPTX reports. 620+ campaigns processed.

n8nClaude APIpython-pptxETL
03

Telegram Bot Ecosystem

10+ production bots: RAG-powered AI assistant, lead capture from chats, CRM notifications, infrastructure monitoring, MTProto channel parsing.

Telegram APIMTProtoClauden8n
04

Infrastructure Watchdog

40+ Docker container monitoring on 2 servers: health checks every 5 min, auto-restarts, Telegram alerts with diagnostics, uptime dashboard.

n8nDocker APIUptime KumaTelegram
05

RAG Knowledge Base

Corporate knowledge base on Dify + Weaviate: 1000+ document indexing, vector search, natural language queries, 11 Docker containers.

Dify 1.12WeaviateRAGDocker
06

Multi-Source Aggregation

ETL from 4+ sources: HH.ru API, Kontur.Zakupki, B2B-Center SOAP, Telegram MTProto. Deduplication, AI scoring, data enrichment.

FastAPISOAPTelethonClaude
Blog

Articles & notes

May 2026

AI Week, May 14–15: the state reaches toward the frontier, users keep breaking chatbots

Two events frame the week: CAISI agreements for pre-deployment testing of all five frontier AI labs, and a viral 18,000-cup water order at a Taco Bell drive-thru as a way to bypass voice AI. Top-down state control and bottom-up user pushback collided in the same week.

Read more

The past week brought a rare mix of events at very different scales — yet with a common thread: who keeps AI under control, and how. Top-down, state regulation reached all the leading labs. Bottom-up, users keep finding ways to work around AI agents in customer-facing products.

On May 5, 2026, the U.S. Department of Commerce, via CAISI (Center for AI Standards and Innovation at NIST), announced agreements with Google DeepMind, Microsoft, and xAI for pre-deployment testing of their models. Together with existing 2024 agreements with OpenAI and Anthropic, all five leading frontier labs are now under a single pre-release evaluation process. Tests run in classified government infrastructure on models with safeguards turned off — meaning what's being tested is raw capability, not a polished consumer product. The trigger is obvious: Claude Mythos and Anthropic's restraint around it. Treasury Secretary Bessent specifically convened the heads of America's largest banks to discuss cyber risks, and the White House built the process directly in response.

The AI Doomsday Clock — our channel rubric — moved back from 11:53 to 11:51. Two minutes, not five: the agreements are formally voluntary with no sanctions; the Commerce Department quietly removed program details from its site a week later; CAISI never receives model weights. The institutional mechanism exists, but it has no teeth yet.

At the other end of the spectrum is Taco Bell's drive-thru AI. Voice ordering by Omilia was deployed at 500+ restaurants and, by spring 2026, more than 890. A California customer ordered 18,000 cups of water to bypass the bot and reach a human. TikTok caught fire within two days. By August 2025, Taco Bell announced it was «reevaluating» the approach — yet by April 2026 it had expanded the rollout by hundreds of locations. Reevaluation continues mid-flight.

What ties the two events together. In both cases, the AI is optimized for its primary metric — capability for the frontier model, accepted orders for the customer-facing agent. Verification, brakes, escape hatches — that's infrastructure that has to be built around them. The state has CAISI (imperfect, but it exists). Most businesses deploying customer-facing AI simply don't.

Two layers of AI control are taking shape across the industry. At the frontier-model level, governmental (voluntary today, mandatory tomorrow). At the customer-facing-product level, engineering (output filters, escape hatches, out-of-domain detection). Between them sits a layer no one is closing yet: middleware products where AI handles real business decisions but doesn't touch frontier capability. Next year's precedents will be carved out in exactly this space.

If your AI agent is alone on shift and humans only enter the picture at an 18,000-water-cup order, TikTok will catch you within weeks. If your frontier model is the kind the Treasury Secretary specifically gathers bankers about, welcome to government pre-deployment review. Between those two extremes lies all the practical architecture work for AI systems in 2026.

May 2026

AI Week, Apr 29 – May 6: launching the «AI Doomsday» channel

Five posts in eight days: channel manifesto, two AI fail breakdowns, a legal-tech case, and the first backward movement of the clock. Summary of week one.

Read more

In late April I launched the «AI Doomsday» channel — a coordinate system for assessing how close the industry is to the point where AI development slips out of control. The metaphor is borrowed from the Bulletin of the Atomic Scientists, who have run the Doomsday Clock for nuclear threat since 1947. Starting position: 11:55. Every major AI event moves the hand with reasoning — or it doesn't.

The manifesto established the frame. April 2026 compressed into one month what used to stretch over years — releases of GPT-5.5, Grok 5, Claude Mythos. Mythos turned out to be a special case: for the first time in its history, Anthropic refused to release a model publicly, granting access to eleven organizations via Project Glasswing to hunt vulnerabilities. Two weeks later Google answered with the opposite doctrine — a universal Gemini 3.1 Pro plus a fleet of security agents. Two strategies for one threat: one lab closes the model, two others wrap it in control infrastructure.

The Dragunsky fail (April 29) is a textbook example of what happens when AI filtering operates on substrings. Eksmo enabled AI screening of manuscripts for «drug propaganda» under the law that took effect on March 1, 2026. Three weeks later the model flagged writer Denis Dragunsky's surname — because «драг» matched the English «drug». Pushkin, Gogol, Tolstoy, and a Bulgakov biography fell under the same filter. Error type — lexical match without semantic understanding, recall-over-precision optimization at maximum sensitivity.

Legal AI as a product class crossed the point where automation savings are zeroed out by the cost of errors (May 4). In the first months of 2026, U.S. courts imposed 145,000 dollars in fines for hallucinated AI citations in legal filings. Escalation by the quarter: 2,500 in January, 7,500 in March, 30,000 in early April, 110,000 in Oregon on April 4, and Greg Lake's indefinite suspension from practice on April 16 in Nebraska. Meanwhile, 61% of federal judges use AI themselves. The systemic problem is in the deal architecture, not in adaptation speed.

The clock moved backward for the first time on May 5 — from 11:55 to 11:53. Reason: mechanistic interpretability has for the first time crossed from academic niche into engineering practice. MIT Technology Review listed it among the 10 breakthrough technologies of 2026, ICLR 2026 in Rio held a dedicated workshop, and on April 30 the first startup released a public LLM debugging tool. Anthropic's Microscope decomposes the model's activation superposition into interpretable features — you can see what exactly happens before a hallucination or jailbreak. Two-minute shift, not five — because the tool is confined to Anthropic and the industry hasn't matched it yet.

The discount fail on May 6 — an English online store is legally bound to fulfill an 8,000-pound order at 80% off because the chatbot promised it at 5 AM. An hour of flattering questions, gradual coupon escalation from 10% to 80%, a fake code in the order comments. Under UK consumer law, the business is liable for AI promises as for those of a rogue employee. Error type — classic prompt injection via social engineering: the bot was designed with an open scope of competence. Per IEEE S&P 2026, 13 percent of e-commerce sites run chatbots with the same open architecture.

What I take away from the channel's first week. Mythos and Microscope from Anthropic are the only strong movements in the right direction. The rest of the industry keeps shipping the maximum possible. Regulation is fragmented and reactive. Financial penalties for careless AI use in legal-tech already grow faster than savings from AI itself; healthcare is next. Five minutes to midnight — a compromise between «AGI is around the corner» (no, it isn't) and «everything is under control» (no, it isn't).

February 2026

Vibe Coding: from marketer to AI engineer

How AI-assisted development enables building production apps without a traditional CS degree.

Read more

It all started with chatbots. While setting them up for business tasks, I kept hitting limitations: bots lacked autonomy, context, and decision-making ability. I wanted systems that could act independently, not just answer questions. That's how I dove into automation, neural networks, and eventually arrived at what's now called vibe coding.

For me, vibe coding isn't just "asking AI to write code." It's a chain of interconnected processes: first visualizing the end product, then decomposing it into dozens of small tasks, writing clear instructions for the AI agent, the "magical" code generation process itself, testing, optimization, and finally — launching to production. Each step requires understanding architecture, business logic, and user experience.

My first projects were classic: document workflow automation, marketing analytics collection — tasks I understood well from 15 years in marketing. But complexity grew: an app for a cosmetologist with YClients sync, a lead generation platform with AI scoring, email campaign analytics with GPT-4 Vision. Each project was more ambitious than the last.

The main lesson — practice, practice, and more practice. Don't mindlessly repeat everything you see on social media. It's far more effective to come up with your own project that solves a real problem and bring it to production. Real understanding comes through deployment, debugging, and scaling.

This path is for anyone willing to understand how LLM models and agent systems work. For those who crave new information and are ready to absorb it in large volumes literally every day. A CS degree isn't required — but curiosity and persistence absolutely are.

August 2026

Self-hosted infrastructure for AI development

UGREEN NAS + two VPS + a self-hosted network coordinator: building production infrastructure for AI projects on your own hardware — and what to do when a protocol stops working.

Read more

At first, like everyone, I used the cloud. But as projects grew more complex, server costs kept rising while resources became insufficient. At some point it became obvious: serious AI development needs its own infrastructure. That's how UGREEN DXP4800+ appeared — a NAS with Intel Pentium Gold 8505, 64 GB DDR5, and a 1.8 TB SSD for Docker containers.

The architecture ended up three-tiered. The NAS runs 75+ Docker containers: Claude Code development environment, CosmoDoc, Plant Care, fit-tracker, n8n automations, Dify with RAG pipelines, Ollama for local LLMs, Jellyfin, Immich, and the full media stack. An Amsterdam VPS serves as the web front: Nginx with Let's Encrypt serves a dozen subdomains, and SSH reverse tunnels forward ports from the NAS to it — public access without a single port forward on the home router. The third point is a small VPS in St. Petersburg running a self-hosted private-network coordinator (Headscale) and a DERP relay.

The most expensive lesson was about protocols. For six months the front ran on VLESS Reality, disguising VPN traffic as ordinary HTTPS on port 443. That stopped working in spring: carrier-side filtering started flagging IPs by TLS fingerprint and throttling them to 500 bytes/s. Neither a new server nor NaiveProxy helped — any TCP+TLS traffic on a flagged address died the same way. The fix was moving to UDP: obfuscated AmneziaWG plus Hysteria 2 over QUIC. The second lesson from the same episode: test a new server from a mobile network before migrating services onto it. One rented VPS turned out to be fully blocked by a mobile carrier — invisible over wired internet.

A separate challenge — unblocking the services themselves. Not only Anthropic and OpenAI APIs are blocked from Russia, but also Cloudflare proxy, TMDB, and servarr.com. The solution is a proxy container routing blocked traffic through the Amsterdam VPS. Cloudflare, meanwhile, only works in DNS-only mode. Every new service requires checking: does it need a proxy?

Among the pitfalls: UGOS kernel ACL blocks shared folder access for non-root users, a custom resolv.conf breaks Docker DNS, Docker on the NAS ran out of subnet pools, and an SSH tunnel with ExitOnForwardFailure dies entirely if a single forwarded port is taken. Each problem is experience that saves hours in future projects. Self-hosted isn't just about cost savings — it's full control over data, performance, and architecture.

February 2026

Multi-agent AI systems in practice

Deploying Dify Platform: RAG pipelines, agent orchestration, and the context problem in team collaboration.

Read more

My main experience with multi-agent systems comes from deploying Dify Platform on self-hosted infrastructure. 11 Docker containers: Flask API server, Next.js frontend, PostgreSQL, Redis, Weaviate for vector search, a sandbox for safe code execution, Celery workers for background tasks. The visual workflow builder lets you chain LLM calls, RAG queries, and custom tools without writing code.

The toolkit includes Claude Agent SDK for programmatic orchestration, MCP servers for connecting external data sources, and Dify itself as a platform for visual AI application design. Each tool fills its niche: SDK for custom logic, MCP for integrations, Dify for rapid prototyping and team collaboration.

The main challenge I faced — context preservation. When a multi-agent system is used by a team, it's critical that each member has access to colleagues' query context and can retrieve it. RAG pipelines with Weaviate partially solve this: documents, conversations, and previous query results are indexed and available to all agents. But a complete solution requires thoughtful memory architecture — short-term, long-term, and episodic.

Where is the technology heading? It's impossible to predict the scale of changes. Perhaps we're on the verge of fully autonomous multi-agent systems that handle the entire development cycle independently — from requirements analysis to deployment — only requesting final approval from the developer to launch the finished product. Agents already write code, test it, and deploy. The only question is when the level of trust will allow removing humans from the loop.

Let's work
together

Open to offers — full-time, project work, or AI automation consulting.

Following the AI industry, writing regularly