01 · sites & platforms
A website that works for you.
E-commerce, bookings, client areas. Or the rescue of a site that no longer works.
available for new projects
Websites, automation and AI agents for your business
less manual work, more time and more clients.
Alberto Di Maria
AI product engineer · Milan
live demo
Hi, I’m D10S. Tell me what you need.
01what I do for you
01 · sites & platforms
E-commerce, bookings, client areas. Or the rescue of a site that no longer works.
02 · automation & AI
Invoices, emails and calendars that talk to each other: repetitive work closes itself.
03 · voice agents
A 24/7 phone assistant that books, reschedules and informs.
how we work
free, to understand the problem
fixed price, agreed upfront, no surprises
demos and updates along the way
training on use, and support after release
02Selected work
A musician’s broken website rebuilt from scratch — fast, secure, trilingual, with a paid client area.
Rebuilt as a modern app — public site + private client area, in three languages (IT/EN/ES) — from a site that barely worked.
Every project gets its own room — versioned audio, comments pinned to the exact second, downloads reserved for paying clients.
Delivered hardened — automatic checks on every change, secured, anti-spam contact form; two access leaks found and closed.
ownership + is_paid RLS paywall · signed-URL downloads · GitHub Actions typecheck/lint/build/test gate
A language tutor you talk to — an Android app that listens, corrects, and remembers.
You learn by speaking — a real conversation, out loud, in English or Spanish — from minute one. Almost no text, by design.
It remembers you — every call is mined for the mistakes you actually made; spaced repetition resurfaces the weak ones and picks tomorrow’s scenario.
Practice built from your errors — writing, translation and vocabulary drills are generated from the mistakes you actually made — graded by an LLM, and only the confident calls feed your memory.
Four agents, not one prompt — absolute beginners get an agent of their own — slower voice, listen-and-repeat, no corrections.
a per-user memory of the mistakes you make drives every scenario and drill — no vector store, just structured recall of your own errors · text drills graded by an LLM, only high-confidence results written back so the memory stays clean · an append-only ledger meters every second of voice and cannot double-spend · prompt-injection defence on user text entering the agent prompt (rejected 6/6) · native criteria score real conversations in production, not just simulations
Lyrics-first music discovery — a playlist that travels your emotions.
Describe a mood — in plain language or tap a 12-emotion wheel — a Claude agent resolves it and Lyra builds the playlist as a trajectory through that space, not a flat list.
Chosen by the lyrics — sentence-transformer embeddings place each track on a valence × energy taxonomy; the exact cited line that matched your mood is surfaced via Musixmatch.
Steer live — more like this / change the mood / raise the energy — reshaping the queue without cutting the current track; a 3D compass turns to your mood and traces the path.
A producer’s online store: beats for sale with previews, licensing and in-page checkout.
In-page checkout — waveform previews, a multi-license cart (MP3 / WAV / stems / exclusive) and Stripe without leaving the page.
Delivers itself — every purchase sends private download links, valid only for the buyer.
Running costs cut to the bone — audio flies from Cloudflare straight to the browser, no server in between.
Vitest + Playwright · GitHub Actions blocking merges · RLS default-deny
An AI agent that answers questions about a company’s real data — exact numbers, nothing made up.
Answers from real data — ask in plain language about clients, orders or calls — the agent reads the CRM, the ERP and the documents and cites its sources.
Math done by code, not the AI — the model only picks where to look and how to explain; the numbers are always exact.
Top 10 at Cursor × Yellow Tech — qualified for the Italian National Hackathon League.
A dance school’s AI phone assistant: answers 24/7, books, reschedules, informs.
Handles the call on its own — recognises the student, gives course info, books, cancels and reschedules — and hands to a human when needed.
Italian and Spanish, no talk-over — answers instantly; confirmations and reminders go out by SMS.
Tested before going live — 9/9 real scenarios passed, with a panel tracking the time and cost of every call.
fixed-scenario eval runner on the live LLM + DB · per-turn token/cost metrics · /observability + /evals dashboards
Autonomous voice agent with outbound calling.
Collects the booking in-browser — chats with the user to gather the details before calling.
Calls the restaurant itself — autonomously dials via Twilio and handles the full phone conversation end-to-end.
Remembers you — persistent cross-session memory that survives ElevenLabs’ stateless calls — recognises returning customers and proposes “the usual”.
Admin dashboard — session analytics and revenue tracking.
ground-truth extraction evals · ElevenLabs online eval criteria · /observability latency dashboard
AI publishing pipeline for Beat Store.
Headless CLI — point it at a producer’s drop folder and it publishes each beat to Beat Store through the store’s authenticated API.
AI metadata + SEO — enrichment via Claude API; idempotent, resumable upload pipeline.
You confirm before publish — it reviews every drop in the terminal and asks first.
Audio generation via RAVE latent-space interpolation.
Interpolate in latent space — upload up to 4 audio files and blend their RAVE latent encodings via a 2D board.
Barycentric weighting — click position sets the blend across all inputs simultaneously.
FastAPI + TorchScript — backend on HF Spaces, React frontend on Vercel.
MSc thesis: from audio analysis to music generation.
Analysis → generation — a three-client pipeline: 3D mood visualization → emotion mapping → prompt-free AI generation, looping in real time.
Sacred geometry as interface — Metatron’s Cube and Platonic solids map emotions to audio features in an interactive 3D client.
Prompt-free generation — via Suno API; comparative benchmark of generative models (Suno, RAVE, MusicGen, Jukebox).
Analog circuit modeling on dedicated effects hardware.
First WDF on the H9000 — a Wave Digital Filter algorithm modeling linear and nonlinear circuits directly on the hardware via VSig3.
Nonlinearities via CPWL — a Canonical PieceWise-Linear representation; validated end-to-end with a diode-clipper circuit.
No prior art — built within VSig3’s low-level limits with no examples in the literature.
Teaching-management automation system.
Automates invoicing — via Fiscozen, and sends personalized client emails.
Syncs & books — Google Calendar sessions → Notion; auto-books on third-party platforms.
Runs daily — across invoicing, email dispatch, session logging and booking.
Dual voice-agent airline support with live escalation — Yellow Tech × ElevenLabs.
Two agents, live escalation — Aria handles passengers on the front line and escalates complex cases to Marco, a supervisor agent.
Multilingual — conversational support (IT / EN / ES / FR) on ElevenLabs Conversational AI.
Hackathon build — made at Yellow Tech × ElevenLabs.
Deep-learning music genre classification (NN / CNN / RNN-LSTM).
Genre from MFCCs — classifies tracks with librosa features, trained on the GTZAN dataset.
NN vs CNN vs RNN-LSTM — compared with accuracy/loss evaluation; scalable to other datasets and models.
Interactive installation on the eco-impact of daily actions.
Act, see the impact — users mimic everyday actions and watch their CO₂ and climate impact in real time on a responsive 3D globe.
Gesture → visuals → sound — Python gesture recognition + TouchDesigner real-time 3D + SuperCollider generative soundscapes.
03About
Engineer (MSc, Politecnico di Milano). I build websites, automation and AI agents for businesses and small companies — from the first line of code to launch, and I keep looking after them afterwards.
I follow what comes out as soon as it does, and I try it: workshops, hackathons, continuous training. Freelance in Milan, working remotely too.
End-to-end AI web tools for the music industry: e-commerce, automation agents and voice interfaces.
Music information retrieval, 3D interfaces and generative audio; full-stack prototype in React, Flask and Python.
Built a custom LLM agent harness from scratch: tool-calling loop, parallel subagents, RAG vs grep over a Markdown knowledge base. No frameworks.
GenAI for music creation with team-based project development.
04Contact
Let's work together.
Write me a couple of lines about what you need. I’ll reply with a first idea of how I’d approach it.
I usually reply within a day.