0
I Use This!
Moderate Activity
Analyzed 2 days ago. based on code collected 2 days ago.

Project Summary

Declarative large-scale machine learning (ML) that aims at flexible specification of ML algorithms and automatic generation of hybrid runtime plans ranging from single-node, in-memory computations, to distributed computations on Apache Hadoop and Apache Spark.
ML algorithms are expressed in an R-like or Python-like syntax that includes linear algebra primitives, statistical functions, and ML-specific constructs. This high-level language significantly increases the productivity of data scientists as it provides (1) full flexibility in expressing custom analytics, and (2) data independence from the underlying input formats and physical data representations. Automatic optimization according to data and cluster characteristics ensures both efficiency and scalability.

Tags

cluster distributed dml hadoop java machine_learning pydml python spark

Apache License 2.0
Permitted

Commercial Use

Modify

Distribute

Place Warranty

Sub-License

Private Use

Use Patent Claims

Forbidden

Hold Liable

Use Trademarks

Required

Include Copyright

State Changes

Include License

Include Notice

These details are provided for information only. No information here is legal advice and should not be used as such.

This Project has No vulnerabilities Reported Against it

Did You Know...

  • ...
    in 2016, 47% of companies did not have formal process in place to track OS code
  • ...
    data presented on the Open Hub is available through our API
  • ...
    use of OSS increased in 65% of companies in 2016
  • ...
    you can embed statistics from Open Hub on your site

Languages

HTML
59%
Java
30%
JavaScript
6%
14 Other
5%

30 Day Summary

Jul 6 2026 — Aug 5 2026

12 Month Summary

Aug 5 2025 — Aug 5 2026
  • 244 Commits
    Down -142 (36%) from previous 12 months
  • 44 Contributors
    Up + 3 (7%) from previous 12 months