Relari (YC W24)AI araçları için test, değerlendirmesi ve sentetik veri oluşturma platformu.
Genel Bakış
Temel özellikler
- Sentetik veri kümesi oluşturma
- Otomatik agent değerlendirme.pipeline'leri
- Senaryo ve konuşma simülasyonu
- Özenli değerlendirme ölçütleri
- Geliştirilmiş LLM uygulamaları için gerileme testleri
- Sürat performansı ve raporlama için benchmarking
Fiyatlar
- Model
- Free
- Kategori
- İzlenebilirlik
- Puan
- 4.3 / 5 (6)
Kullanım senaryoları
AI araçlarının test edilmesi
Güvenilir ve test edilmüş AI araçları değerlendirmesi için.
Artılar ve eksiler
Artılar
- Çok adımlı AI araçları için özel olarak tasarlandı.
- Düzenli sentetik test verisi üretmek için ölçekli.
- Özel ölçütler ve değerlendiriciler desteklenmektedir.
- Y Combinator tarafından desteklenmekte ve aktif geliştirme sürdürülmektedir.
Eksiler
- Primarily ana olarak teknik takımlar için yönlendirilmiştir, geliştiriciler olmayan için değil.
- Yenilikçi bir platform olduğu için özellikleri bir evrim geçiriyor.
- Varolan stacklar için uyum için entegrasyon çalışmaları gerekebilir.
Savaş rekoru
Pantheon’da 3 savaş.
Last 3 battles
İncelemeler
6 puandan ortalama.
İnceleme bırakmak için giriş yap.
Skeptical, then convinced
I went in skeptical — most tools in this space overpromise. It actually delivers on customizable evaluation metrics, and purpose-built for evaluating multi-step AI agents caught me off guard. still, I'd recommend giving it a real trial.
Solid for our team
We rolled this out across the team last quarter and supports custom metrics and evaluators. Customizable evaluation metrics fits neatly into how we already work, and customizable evaluation metrics removed a step we used to do by hand. Primarily aimed at technical teams, not non-developers, which is the main caveat, but it has held up under daily use.
Skeptical, then convinced
I went in skeptical — most tools in this space overpromise. It actually delivers on performance benchmarking and reporting, and supports custom metrics and evaluators caught me off guard. Primarily aimed at technical teams, not non-developers is why this isn't a perfect score, still, I'd recommend giving it a real trial.
Compared a few options
Evaluated this against two competitors. Where it wins: scenario and conversation simulation and purpose-built for evaluating multi-step AI agents. Where it lags: may require integration work to fit existing stacks. On balance the feature set — especially scenario and conversation simulation — justifies the 5 stars for our use case.
Use it every day
Honestly didn't expect to like it this much. Performance benchmarking and reporting is exactly what I needed, and purpose-built for evaluating multi-step AI agents. I do wish may require integration work to fit existing stacks, but I reach for it almost every day now and it just clicks.
Skeptical, then convinced
I went in skeptical — most tools in this space overpromise. It actually delivers on regression testing for LLM apps, and supports custom metrics and evaluators caught me off guard. May require integration work to fit existing stacks is why this isn't a perfect score, still, I'd recommend giving it a real trial.
Sorular
What are the pros of using Relari?
Relari is purpose-built for evaluating multi-step AI agents, generates synthetic test data at scale, and supports custom metrics and evaluators, with active development backed by Y Combinator.
Asked by Anya Sokolova · Jun 27, 2026
Is Relari suitable for non-technical teams?
Relari is primarily aimed at technical teams, not non-developers, and may require integration work to fit existing stacks.
Asked by Hasan Demir · May 27, 2026
What features does Relari support?
Relari supports features like synthetic dataset generation, automated agent evaluation pipelines, scenario simulation, and customizable evaluation metrics.
Asked by Marisol Pena · Apr 17, 2026
What is Relari used for?
Relari is a platform for testing, evaluation, and synthetic data generation for AI agents, helping teams improve their reliability through systematic testing and evaluation.
Asked by Hana Kobayashi · Mar 25, 2026
Soru sor
İzlenebilirlik alternatifleri

Geliştirici platformu için bir araya gelen LLM uygulamalarını inşa etmek, izlemek ve ölçeklendirmek.

Otonom AI agentleri ve zeki sistemler için güvenlik ve yönetim platformu.

Sonuçtan sonuç-a platform için AI agent'lerin değerlendirmesi, izlenmesi ve geliştirilmesi

Size ait markanızın ChatGPT, Claude, Perplexite, ve Google AI Özetleri üzerinden nasıl gösterildiği izlenir.

Code-less bir AI akış oluşturucu aracın, çoklu büyük dil modellerini (LLM) entegre ederek işlemleri otomatize etmek için işletmelerde işlem yapmalarını sağlar. İpuçlarını bağladığında, birçok akışa kolayca ve code-less katılması sağlar.

İşlem otomasyonu için AI ajansı inşa, değerlendir ve iyileştir
Tamamı gözden geçirilebilir platform, üretim LLM uygulamalarını izlemek, hata ayıklamak ve geliştirmek için.

İT operasyonları için çalışan bir AI ajansı, ihbarda deteksiyon, triaj ve çözüme gecikmeyi önler.
Trending now

Belge zeka API, karmaşık PDF'ler, sunumlar ve tabloları okur ve yapılandırılmış verileri ayırarak, ayrıştırır, ayırır, OCR yapar ve çıkartır.

Sponsorlu yanıtlar, tık başına tahsil edilen

Daha doğru Homework Yardımıyla Tam Açıklamalar

Çıkarılmış modda 12B model, 128K bağlam penceresi ile karışık görüntü ve metin ile işlenir.
