This one is closed
Live roles like this one
-
F
8h ago
Salesforce Technical Architect
Alithya United States
-
D
12h ago
Scalian Spain
- C 21h ago
- C 1d ago
See every "Location Pisa Italy" role →
Get new “Location Pisa Italy” roles by email
One email a day with what is new in "Location Pisa Italy". Nothing new, no email.
We confirm the address first, and every mail carries an unsubscribe link. Alerts are ours, not a third party's.
Why this grade
This listing scored 58/100, which is a C. It lost the most ground on pay transparency.
- Description depth 20 / 20 How much the posting actually says about the work, measured in characters of real text.
- Freshness 15 / 15 How recently it was posted. Older postings are likelier to be filled or abandoned.
- Remote clarity 15 / 15 Whether "remote" means anywhere, or is quietly restricted to one country.
- Corroboration 5 / 10 Whether more than one source carries this listing.
- Role specificity 3 / 10 Whether the listing is tagged well enough to tell what the role actually is.
- Pay transparency 0 / 25 A published salary range, worth more than any other single factor because it is what a candidate cannot find out without applying.
Every figure above is arithmetic over the posting itself — its salary field, its text, its age, its tags and how many sources carry it. How the grades work →
This listing does not state a salary
$140k – $226k
That is the middle half of what comparable roles paid on this board over the last 90 days — 30 listings that did publish a figure, median $188k. It is not this employer's offer, and we have no idea what they pay. It is only what the rest of the market advertised.
hn-hiring
Remote: Yes, remote only
Willing to relocate: No
Technologies: Python, LLM/agent orchestration, local embeddings, retrieval evaluation, AWS (Lambda/SQS/EventBridge), Terraform, Django, PostgreSQL, LightGBM/PyTorch, NLP/Transformers
Résumé/CV: linkedin.com/in/vslovik
Code: github.com/vslovik/fenix — local-embedding search whose relevance is actually measured: labelled control probes, blinded human ranking, precision@k. No API keys.
Email: [email protected]
Software architect, 15+ years in production systems, almost entirely startups and internal startups — fintech, e-commerce, pharma, publishing.
The work I get pulled into is the recurring startup problem: a service shipped fast under launch pressure, without adequate tests, that later has to be made reliable without being stopped. Incident response, re-architecture, and the release discipline that keeps it from happening again. Most recently that has meant a regulated UK consumer-credit platform — loan servicing, arrears, forbearance, statutory breathing space, and early-settlement calculations written against consumer-credit legislation. Regulation as code, behind a test suite larger than the production codebase.
I've done that in all three configurations: taking a core system from problem statement to release, leading the team that carried it (1 to 7 engineers in ten months), and now doing the same work again with agentic tooling covering what the team used to.
On the data side: a LightGBM acquisition model over a 38M-row base — 0.77 test AUC, 8x lift in the top 1% — scoring 2.9M households for a live campaign. The part I'd rather be judged on is what happened next: I found a validation-set defect in my own pipeline (early stopping on the test split), quantified its effect across every figure I had already reported, restated them, and added a pure-noise regression test that pins the model to chance when fed random features — so that class of leak cannot come back quietly. NLP is hands-on rather than API-deep: my degree thesis fine-tuned BERT, RoBERTa and XLNet to state of the art on the FNC-1 stance-detection benchmark, published at LREC 2020.
Building on my own time: github.com/vslovik/fenix — it ranks an incoming stream against a free-text description of what you're looking for, and answers questions over the same corpus with citations back to source chunks. Ollama embeddings, sqlite-vec, no API keys. The part worth looking at is the evaluation: the ranking anchor is scored against a labelled probe set with a deliberate control group of things I don't want, and live results are rated blind — scores hidden, order shuffled — so the human judgement stays independent of the ranking it is judging. Doing that produced a measured finding I did not expect: an embedding has no notion of negation, so naming a technology in order to reject it moves the anchor toward it. Numbers and method in lessons/embedding-anchors.md.
Also a tool-calling agent that turns unstructured regulatory text into a deterministic calculation pipeline — the model does the extraction, a deterministic engine does the arithmetic.
Looking for agentic AI/LLM engineering, LLM evaluation and observability, AI integration, or software architecture. Founding-engineer shape suits me — early employee, not co-founder, but early enough to be in the room where the work gets defined. Direct with the company that owns the product: not consultancy, not agency placement, not a body on someone else's engagement.
Quick question · anonymous · one tap
Would you apply to this job?
Answer to see what other job seekers said.
Your turn · no account needed
Help the next applicant
You may know something about this listing that we cannot see from here. One tap. No account needed. Signed-in reports earn points once the evidence agrees with you.
I know what it pays
Sign in with Google to earn points for reports — 100 confirmed points buy a week of Early Access.
Where this listing came from
- 01 Oct 2026 Hacker News first sighting
Seen on 1 board over 0 days.