Computer Vision
Production-grade vision for medical, industrial and earth-observation. Detection, segmentation, super-resolution and real-time inference on the edge.
We are an AI research and engineering studio for tier-1 banks, insurers and industrial leaders. Built by scientists with world-record inference, published medical research, and operator experience at NVIDIA, Siemens Healthineers and JP Morgan — not prompt engineers.
DeepSeek V3 inference throughput — open-source, reproducible.
A studio, not an agency.
Scientists, not slideware.
Archeon is a small, senior team of French and American AI scientists based in San Francisco, Zürich, Paris, and Dubai. We embed with your engineers, ship working systems, transfer knowledge — and leave behind infrastructure your team owns.
DeepSeek V3 inference throughput — published, reproducible.
Medical imaging research with Siemens Healthineers and UCL.
Built systems at NVIDIA, JP Morgan, Siemens Healthineers, Avian.
Five engineering practices, one team. We sequence them per engagement — vision in production, language at scale, ML infra that survives audit.
Production-grade vision for medical, industrial and earth-observation. Detection, segmentation, super-resolution and real-time inference on the edge.
On-prem and air-gapped LLMs. Custom kernels, quantization, speculative decoding — the work that took us to a world record on DeepSeek V3.
Clinical-grade models with regulatory awareness. Surgical assistance, imaging and diagnostics that meet hospital reliability bars.
Risk, alpha research and document intelligence for banks and asset managers. Models that survive compliance, audit and stress.
Frontier research turned into deployable systems. We co-author with Berkeley, Harvard and UCL — and ship what works to your stack.
We work where mistakes are expensive — regulated industries with measurable AI ROI. Pick a vertical to see how we engage.
We build risk engines, on-prem LLM stacks and document-intelligence systems for tier-1 banks and asset managers. Auditable from training data to inference, deployed inside your perimeter — no third-party model APIs.
Engagements where the difference between research and production was measurable — in latency, in dollars, in patient outcomes.
Our team built and ran AI inside the institutions our clients trust. We come back with the playbook.
AI & Accelerated Computing
Inference optimization, custom CUDA kernels and accelerated training pipelines.
Medical AI
Clinical imaging, diagnostic models and registration systems shipped into hospitals.
Financial AI
On-prem LLM tooling and document intelligence inside a tier-1 banking perimeter.
Archeon is a small, senior team of AI researchers and engineers based in San Francisco, Zürich, Paris, and Dubai. We embed with your operators, ship working systems, and leave behind infrastructure and people your team owns.
Senior practitioners, no junior consultants. The names on the proposal are the people in your meetings.
LLM Engineering
Production ML across healthcare and fintech. World-record LLM inference at 303 tokens/sec on DeepSeek R1. ML research on diffusion models for CT imaging at Siemens Healthineers.
Computer Vision Engineering
Ex-UC Berkeley research, published in astronomical ML. Edge OCR architect — 10× faster inference on CPU-only hardware. Real-time medical ultrasound AI and generative vision from R&D to production.
Medical ML Engineering
Doctoral researcher at UCL. ML engineer on radiotherapy optimization and medical image registration at Siemens Healthineers. LLM inference specialist with CUDA, TensorRT, and speculative decoding.
Geospatial AI & Remote Sensing
Deep learning engineer leading applied research and performance benchmarking. Computer vision on satellite imagery for geospatial asset tracking. Former ML engineer at Upstage building LLM systems.
Active collaborations with research groups at Berkeley, Harvard and UCL.
“Archeon's hyperspectral analysis pipeline transformed our environmental monitoring program. Their team delivered accuracy levels we didn't think were achievable with satellite data alone.”
Bring a real problem. We bring a senior engineer and a scientist. 30 minutes — no slides, no qualifying call.