# 使用 Homebrew 安装 alluxio

查看 alluxio 的安装路径、可执行文件、元数据以及面向 AI 代理工作流的安全说明。

## 安装

```sh
sudo av install brew:alluxio
```

其他安装命令:

### macOS

- Homebrew (100%):

```sh
brew install alluxio
```

  证据: local Homebrew formula metadata

## 软件包事实

- **软件包键:** brew:alluxio
- **软件包管理器:** Homebrew
- **版本:** 2.9.5
- **来源摘要:** Open Source Memory Speed Virtual Distributed Storage
- **主页:** <https://www.alluxio.io/>
- **已生成:** 2026-08-03T19:37:03+00:00

## 可执行文件

- alluxio (别名)
- alluxio-common.sh (别名)
- alluxio-masters.sh (别名)
- alluxio-monitor.sh (别名)
- alluxio-mount.sh (别名)
- alluxio-start.sh (别名)
- alluxio-stop.sh (别名)
- alluxio-workers.sh (别名)
- launch-process (别名)

## 安装行为

- Bottle: 不可用

## 版本和新鲜度

- 页面生成时间: 2026-08-03
- 管理器版本: 2.9.5
## 项目历史与用法

Alluxio is an open source data orchestration and virtual distributed storage system. It sits between compute frameworks and storage systems, presenting a unified namespace and APIs while caching data close to analytics and AI workloads.

### 项目历史

The Alluxio README states that the project was formerly known as Tachyon and originated as a UC Berkeley AMPLab research project. It was the data layer of the Berkeley Data Analytics Stack, and the README points readers to Haoyuan Li's dissertation, Alluxio: A Virtual Distributed File System, for the research background.

The project later became Alluxio, an Apache-licensed open source system owned by the Alluxio Open Source Foundation and managed by an Alluxio Project Management Committee. Official docs position it as data orchestration technology for analytics and AI, with a memory-first tiered architecture, global namespace, and server-side API translation.

### 采用历史

Official documentation says Alluxio has attracted more than 1000 contributors from over 300 institutions including Alibaba, Baidu, CMU, Google, IBM, Intel, NJU, Red Hat, Tencent, UC Berkeley, and Yahoo. It also says the project is deployed in production by hundreds of organizations, while the README describes petabyte-scale production deployments and a largest deployment exceeding thousands of nodes.

Alluxio adoption followed the data-platform pattern of being useful when compute and storage are decoupled. Its docs call out use with Apache Spark, Presto, Tensorflow, Apache HBase, Apache Hive, and Apache Flink above storage systems such as Amazon S3, Google Cloud Storage, OpenStack Swift, HDFS, IBM Cleversafe, EMC ECS, Ceph, NFS, Minio, and Alibaba OSS.

### 使用方式

Operators deploy an Alluxio master and workers, mount persistent under storage into the Alluxio namespace, and point compute jobs at Alluxio through HDFS-compatible, S3, FUSE/POSIX, REST, or Java APIs. The Spark guide, for example, shows Spark applications reading and writing `alluxio://` paths while Alluxio transparently fetches and caches data from under storage.

The command-line package includes administrative and cluster scripts such as `alluxio`, `alluxio-start.sh`, `alluxio-stop.sh`, `alluxio-masters.sh`, and `alluxio-workers.sh`. In package-manager terms this is not just a CLI utility; it is a local entry point into a distributed service stack.

### 为什么软件包爱好者会关心

Alluxio is significant to package people because it packages a full distributed Java data system behind Unix-style scripts. A Homebrew install can be convenient for development and local testing, but production use usually involves clusters, Docker, Kubernetes, Hadoop/Spark classpaths, and configuration files.

The former Tachyon name, AMPLab origin, and BDAS connection tie Alluxio to the same research-to-production wave that shaped Spark-era big data infrastructure. Its package metadata has to communicate storage, cache, filesystem, and orchestration at once because all of those labels are true in different deployment modes.

### 时间线

- Tachyon era: Project originated at UC Berkeley AMPLab as the data layer of the Berkeley Data Analytics Stack.
- Alluxio rename: The project became Alluxio, described by the README as formerly known as Tachyon.
- Foundation/PMC era: README states the Alluxio Open Source Foundation owns the project and the PMC manages operation.
- Current: Official docs position Alluxio as data orchestration for analytics and AI across many compute engines and storage backends.

### Related projects

- Apache Spark, Presto, Tensorflow, Apache HBase, Apache Hive, and Apache Flink are documented compute-side integrations.
- Amazon S3, Google Cloud Storage, OpenStack Swift, HDFS, IBM Cleversafe, EMC ECS, Ceph, NFS, Minio, and Alibaba OSS are documented storage-side examples.
- The Berkeley Data Analytics Stack and UC Berkeley AMPLab are central to Alluxio's origin story.

### 来源

- <https://documentation.alluxio.io/os-en/api/posix-api.md>
- <https://documentation.alluxio.io/os-en/compute/spark.md>
- <https://documentation.alluxio.io/os-en/overview.md>
- <https://raw.githubusercontent.com/Alluxio/alluxio/master/README.md>


## 安全说明

narrow executable package without higher-risk signals.

- **Geiger 风险:** 绿色 / 低
- narrow executable package without higher-risk signals


## Configuration and credential file locations

These source-backed paths show where this package keeps local settings or durable credentials. Automic Vault can use them as review targets for secret scanning, migration, and command approval.


## Configuration files

- Unix: ${alluxio.conf.dir}/alluxio-site.properties, ${user.home}/.alluxio/alluxio-site.properties, /etc/alluxio/alluxio-site.properties

## Combined YAML source

View the package source record on GitHub. [combined/alluxio.yml](https://github.com/mxcl/pkgdb/blob/main/combined/alluxio.yml)


## 来源

- pkg.so package database
- Geiger risk classifier
- curated configuration and credential file locations
- curated package history
- pkgdb category and tag curation
- cross-ecosystem install command graph
