The Challenge
Das ursprüngliche Hybrid-Setup von Lightricks hielt mit der wachsenden Nachfrage und den immer komplexeren Anforderungen an Machine Learning und Analytics nicht mehr Schritt. In geschäftskritischen Momenten wurde die Infrastruktur zum Nadelöhr: Daten-Uploads stießen an ihre Grenzen, Cluster fielen während Finanzierungsrunden und in Spitzenzeiten aus.
The Solution
Lightricks migrierte zu Google Cloud und setzt für die automatisierte Datenaufnahme und -analyse auf BigQuery und Dataflow. Das Team führte Google Kubernetes Engine für containerisierte Infrastruktur ein und nutzt Vertex AI für Machine-Learning-Modelle. DoiT International liefert dabei laufend Unterstützung – von der Architektur bis zur Problemlösung.
Results
- Rund 1 Milliarde Events pro Tag mit BigQuery und Dataflow verarbeiten
- Funktionsfähige Kubernetes-Infrastruktur auf GKE in wenigen Wochen mit minimalem Engineering-Aufwand aufgebaut
- Infrastruktur-Engpässe beseitigt, die zuvor in geschäftskritischen Phasen zu Systemausfällen führten
- Business Intelligence in Echtzeit für die Optimierung von Werbekampagnen auf Terabytes an Daten ermöglicht
Als wir zum Beispiel einen Cluster auf GKE aufsetzen und an unsere Machine-Learning-Systeme anbinden wollten, hat DoiT dafür gesorgt, dass unsere Data Lakes für die Forschung und unsere On-Premises-Compute sauber zusammenspielen. Unser Machine-Learning-Aufbau ist anspruchsvoll, und DoiT steht uns laufend zur Seite – von der Architektur bis zur Problemlösung.
Ofir Bibi, VP Research, Lightricks
Skalierung als Herausforderung
Mit Apps wie Facetune, Videoleap und Photoleap fand Lightricks schnell den Marktdurchbruch. Doch das rasante Wachstum brachte die hybride Cloud-on-Premises-Infrastruktur an ihre Grenzen. Das System konnte den GPU-Bedarf nicht mehr decken und bildete die zunehmend komplexen Anforderungen an Machine Learning und Analytics nicht mehr ab. Während Finanzierungsrunden und in Spitzenzeiten kam es zu kritischen Ausfällen – eine robustere Plattform musste her.
Datenaufnahme automatisiert im großen Maßstab
BigQuery und Dataflow bilden das Fundament der neuen Datenplattform von Lightricks. Heute laufen rund 10.000 Events pro Sekunde durch das System – täglich eine Milliarde. Die Autoscale-Funktion von Dataflow hat die früheren Engpässe beseitigt, an denen Daten-Uploads regelmäßig an Systemgrenzen scheiterten. Diese Automatisierung macht Business Intelligence in Echtzeit möglich: Teams optimieren Werbekampagnen auf Basis von Terabytes an Daten und sprechen Nutzer aus erfolgreichen Marketingkampagnen sofort gezielt an.
Kleine Teams, große Infrastruktur
Google Kubernetes Engine hat die Infrastruktur-Komplexität für das schlanke Engineering-Team von Lightricks drastisch reduziert. Innerhalb weniger Wochen brachten einige Engineers und DevOps-Mitarbeitende eine funktionsfähige Kubernetes-Infrastruktur an den Start – mit dem Vorgängersystem wäre das undenkbar gewesen. Die Trennung von Storage und Compute hat die Infrastrukturhürden beseitigt, die die Entwicklung zuvor ausgebremst hatten. So können sich die Teams auf den Mehrwert für das Business konzentrieren, statt Systeme zu warten.
Machine Learning auf einem neuen Niveau
Google Cloud hat die Compute-Engpässe für Machine-Learning-Workloads bei Lightricks aufgelöst. Vom experimentellen Cloud-Training im Jahr 2014 hat sich das Team zu einer robusten Plattform weiterentwickelt, auf der Compute-Ressourcen on demand bereitstehen. Marketing, Produktoptimierung und Recommendation Engine entwickeln ihre Machine-Learning-Modelle heute auf Compute Engine – und mit der Migration zu Vertex AI lassen sich Empfehlungssysteme und die Optimierung der Nutzerinteraktion noch schneller skalieren.
Sichere Drittanbieter-Integrationen
Dank der Integrationsmöglichkeiten von Google Cloud kann Lightricks Drittanbieter-Dienste wie Cloudinary und Elasticsearch sicher anbinden. Die Plattform leitet Traffic sicher außerhalb privater Netzwerke weiter, ohne Systeme dem öffentlichen Internet auszusetzen. Dieses Sicherheitsframework stützt die Backend-Entwicklung von Lightricks: Das Unternehmen baut seine Services auf Open-Source-Technologien auf – bei robustem Schutz.
Weiteres Wachstum und Plattformausbau
Für 2022 plant Lightricks einen umfangreichen Ausbau seiner Backend-Services, darunter app-übergreifende Profile und Funktionen für den Medien-Upload. Dieses Wachstum erzeugt mehr Daten und erfordert leistungsfähigere Machine-Learning-Modelle. Mit Google Cloud im Rücken kann das Unternehmen schnell und kosteneffizient skalieren, seine Creator-Plattform weiterentwickeln und dabei State-of-the-Art-Services für die Content-Erstellung liefern.
So unterstützt DoiT Cloud-Teams bei der Kostenkontrolle
Erfahren Sie, wie Cloud Intelligence™ Teams hilft, Transparenz, Governance und Unit Economics in Cloud-Umgebungen zu verbessern.
More customer stories
Finlex halbiert Cloud-Kosten und bringt KI mit DoiT in Produktion
- Over 65%
- Weniger Cloud-Infrastrukturkosten von 2024 bis heute
- 40%
- Kosteneinsparungen durch bessere Transparenz und eine effiziente KI-Architektur
Cloud-Kosten werden bei Hippo zum Hebel für die Teams
- Minutes
- Integrationsdauer von Attribute™ in AWS
- Business unit
- Erreichter Grad an Kostenverantwortung
Island sieht die echten Kosten pro Kunde – ganz ohne Tagging
- $5B
- Unternehmensbewertung
- Days
- Zeit bis zu den ersten Insights
- 0
- Benötigte Tags
Claroty ordnet Kosten dank Attribute™ präzise einzelnen Kunden zu
- 1,000+
- Geschützte Kunden weltweit
- Days
- Zeit bis zu den ersten granularen Kosten-Reports
- Zero
- Beeinträchtigung des Betriebs beim Deployment
Wertbasiertes AI-Pricing dank Kostenattribution auf Kundenebene
- $1.3M
- Umsatz mit negativer Marge aufgedeckt
- ~360
- unprofitable Accounts identifiziert
- $1.3M
- aggregierte Verluste aufgedeckt
Salt Security standardisiert die COGS-Messung für bessere Margen und Preisgestaltung
- 1
- Single Source of Truth für COGS
- 1
- Standardisierte Single Source of Truth für COGS bei CFO und DevOps
- 0
- Code- oder Tagging-Änderungen für die Integration von Attribute™
PropertyGuru macht aus dem Cloud-Kostenchaos echte Engineering-Verantwortung
- Minutes
- um Shared-Service-Rechnungen zu entschlüsseln (zuvor Stunden)
- 32M+
- Monatlich betreute Immobiliensuchende
- 2 weeks
- Vom Rollout bis zu umsetzbaren Erkenntnissen
What they say
We have worked with DoiT for many years, and there has been an increasing number of capabilities and features in DoiT Cloud Intelligence. We've embedded features such as Cloud Analytics and Reports in our own FinOps processes, it's become core to what we do.
Martin Lee, Director of Operations
DoiT gave us the confidence to move from experimentation to production. They helped us understand the right way to build AI for the real world.
Milad Rezazadeh, CTO
Attribute™'s cost grouping technology took our cost visibility and allocation to a whole new level. Now, our teams are fully accountable for their budgets, significantly improving our cloud efficiency and helping us minimize unnecessary costs.
Eli Zilbershtein, Head of DevOps, Hippo
You can't tag a customer in a multi-tenant environment. Attribute™ finally shows us what each customer costs and what's driving those costs.
Omri Cohen, Director of Engineering, Platform
Attribute™'s data is truly unmatched. No other solution on the market could deliver the precise customer cost and usage profiles we needed in such a complex infrastructure. Within weeks, the data from Attribute™ transformed our understanding of cost structures, influencing key strategic decisions in pricing, renegotiations, and market positioning.
Jonathan Langer, COO, Claroty
Attribute™ simplified tracking customer costs in our multi-tenant environments. Customer cost measurement is now clear and standardized, and finance gets the business context they need. Integration was quick and required no changes.
Kfir Lippmann, CFO, Salt Security
Attribute™ translates complex cloud bills into actionable, business-centric insights that empower our engineering teams to take true ownership of their costs.
Balamurugan Mohandossgandhi, Head of IT and Infrastructure, PropertyGuru
This has let us get a better idea of what our cost of goods sold really is. It's not every day you come across something that delivers value as quickly as yours did for us. I was seeing useful insights inside the POC, and we had only deployed it to a couple of real clusters.
Jason Moore, Principal DevOps Engineer, Accrete AI
Eliminating the need to tag thousands of resources has freed up my team and we've invested our efforts in enhancing our platform significantly.
Ziv Sivan, VP of Engineering
PerfectScale by DoiT has become an important part of how we optimize Kubernetes at scale at OneFootball. It gives our platform team the visibility, automation, resiliency insights, and confidence we need to balance cost efficiency with production readiness, especially as we prepare for major global football moments like the 2026 FIFA World Cup.
Andrea Benfatto, Platform/Cloud Runtime Engineering Manager
Cloudflow's new RDS End of Life alerts have allowed us to be more proactive on keeping our database instances up-to-date. The new solution gives us internal visibility ahead of time so that we can prepare for upgrades, instead of having to upgrade under pressure while incurring extended support costs.
Jon Fairbanks, Site Reliability Engineering Manager
PerfectScale cut 40% off our total EKS spend, and the automations handle what used to take our team 20 hours a month. Now we spend that time on reliability and performance instead of chasing cost metrics.
Caio Cristo, Director of Infrastructure/SRE
What I really like about DoiT's approach is that you're very hands-on and proactive. Satyam would ping me a few times a sprint, letting me know about the most current features, checking in on how things are going. When we are going through a peak time, that proactiveness makes a real difference. Satyam always comes through whenever we need support and helps us leverage the right experts to get us where we need to be.
Chiamaka Ibeme, Engineering Manager, Platform
Your cloud bill shouldn't be a mystery
Let us show you what ships this week.

