petals/server2.id at a7f87b636ba136287e60b450d22d452741aefcf3

mirror of https://github.com/bigscience-workshop/petals synced 2024-10-31 09:20:41 +00:00

Alexander Borzunov 8c546d988a

Test Llama, rebalancing, throughput eval, and all CLI scripts (#452 )

This PR extends CI to:

1. Test Llama code using [TinyLlama-v0](https://huggingface.co/Maykeye/TinyLLama-v0).
2. Test rebalancing (sets up a situation where the 1st server needs to change its original position).
3. Check if benchmark scripts run (in case someone breaks its code). Note that the benchmark results are meaningless here (since they're measured on a tiny swarm of CPU servers, with low `--n_steps`).
4. Test `petals.cli.run_dht`.
5. Increase swap space and watch free RAM (a common issue is that actions are cancelled without explanation if there's not enough RAM - so it's a useful reminder + debug tool).
6. Fix flapping tests for bloom-560m by increasing tolerance.

Other minor changes: fix `--help` messages to show defaults, fix docs, tune rebalancing constants.

2023-08-08 19:10:27 +04:00

1.2 KiB

Raw History

View Raw

1.2 KiB Raw History

1.2 KiB

Raw History