ColossalAI/tests/kit/model_zoo/registry.py

#!/usr/bin/env python
from dataclasses import dataclass
from typing import Callable, List, Union

__all__ = ["ModelZooRegistry", "ModelAttribute", "model_zoo"]


@dataclass
class ModelAttribute:
    """
    Attributes of a model.

    Args:
        has_control_flow (bool): Whether the model contains branching in its forward method.
        has_stochastic_depth_prob (bool): Whether the model contains stochastic depth probability. Often seen in the torchvision models.
    """

    has_control_flow: bool = False
    has_stochastic_depth_prob: bool = False


class ModelZooRegistry(dict):
    """
    A registry to map model names to model and data generation functions.
    """

    def register(
        self,
        name: str,
        model_fn: Callable,
        data_gen_fn: Callable,
        output_transform_fn: Callable,
        loss_fn: Callable = None,
        model_attribute: ModelAttribute = None,
    ):
        """
        Register a model and data generation function.

        Examples:

        ```python
        # normal forward workflow
        model = resnet18()
        data = resnet18_data_gen()
        output = model(**data)
        transformed_output = output_transform_fn(output)
        loss = loss_fn(transformed_output)

        # Register
        model_zoo = ModelZooRegistry()
        model_zoo.register('resnet18', resnet18, resnet18_data_gen, output_transform_fn, loss_fn)
        ```

        Args:
            name (str): Name of the model.
            model_fn (Callable): A function that returns a model. **It must not contain any arguments.**
            data_gen_fn (Callable): A function that returns a data sample in the form of Dict. **It must not contain any arguments.**
            output_transform_fn (Callable): A function that transforms the output of the model into Dict.
            loss_fn (Callable): a function to compute the loss from the given output. Defaults to None
            model_attribute (ModelAttribute): Attributes of the model. Defaults to None.
        """
        self[name] = (model_fn, data_gen_fn, output_transform_fn, loss_fn, model_attribute)

    def get_sub_registry(
        self, keyword: Union[str, List[str]], exclude: Union[str, List[str]] = None, allow_empty: bool = False
    ):
        """
        Get a sub registry with models that contain the keyword.

        Args:
            keyword (str): Keyword to filter models.
        """
        new_dict = dict()

        if isinstance(keyword, str):
            keyword_list = [keyword]
        else:
            keyword_list = keyword
        assert isinstance(keyword_list, (list, tuple))

        if exclude is None:
            exclude_keywords = []
        elif isinstance(exclude, str):
            exclude_keywords = [exclude]
        else:
            exclude_keywords = exclude
        assert isinstance(exclude_keywords, (list, tuple))

        for k, v in self.items():
            for kw in keyword_list:
                if kw in k:
                    should_exclude = False
                    for ex_kw in exclude_keywords:
                        if ex_kw in k:
                            should_exclude = True

                    if not should_exclude:
                        new_dict[k] = v

        if not allow_empty:
            assert len(new_dict) > 0, f"No model found with keyword {keyword}"
        return new_dict


model_zoo = ModelZooRegistry()
[test] added timm models to test model zoo (#3129) * [test] added timm models to test model zoo * polish code * polish code * polish code * polish code * polish code 2023-03-14 06:29:18 +00:00			`#!/usr/bin/env python`
			`from dataclasses import dataclass`
[workflow] fixed build CI (#5240) * [workflow] fixed build CI * polish * polish * polish * polish * polish 2024-01-10 14:34:16 +00:00			`from typing import Callable, List, Union`
[test] added timm models to test model zoo (#3129) * [test] added timm models to test model zoo * polish code * polish code * polish code * polish code * polish code 2023-03-14 06:29:18 +00:00
[misc] update pre-commit and run all files (#4752) * [misc] update pre-commit * [misc] run pre-commit * [misc] remove useless configuration files * [misc] ignore cuda for clang-format 2023-09-19 06:20:26 +00:00			`__all__ = ["ModelZooRegistry", "ModelAttribute", "model_zoo"]`
[test] added timm models to test model zoo (#3129) * [test] added timm models to test model zoo * polish code * polish code * polish code * polish code * polish code 2023-03-14 06:29:18 +00:00

			`@dataclass`
			`class ModelAttribute:`
			`"""`
			`Attributes of a model.`
[test] added torchvision models to test model zoo (#3132) * [test] added torchvision models to test model zoo * polish code * polish code * polish code * polish code * polish code * polish code 2023-03-15 02:42:07 +00:00
			`Args:`
			`has_control_flow (bool): Whether the model contains branching in its forward method.`
			`has_stochastic_depth_prob (bool): Whether the model contains stochastic depth probability. Often seen in the torchvision models.`
[test] added timm models to test model zoo (#3129) * [test] added timm models to test model zoo * polish code * polish code * polish code * polish code * polish code 2023-03-14 06:29:18 +00:00			`"""`
[misc] update pre-commit and run all files (#4752) * [misc] update pre-commit * [misc] run pre-commit * [misc] remove useless configuration files * [misc] ignore cuda for clang-format 2023-09-19 06:20:26 +00:00
[test] added timm models to test model zoo (#3129) * [test] added timm models to test model zoo * polish code * polish code * polish code * polish code * polish code 2023-03-14 06:29:18 +00:00			`has_control_flow: bool = False`
[test] added torchvision models to test model zoo (#3132) * [test] added torchvision models to test model zoo * polish code * polish code * polish code * polish code * polish code * polish code 2023-03-15 02:42:07 +00:00			`has_stochastic_depth_prob: bool = False`
[test] added timm models to test model zoo (#3129) * [test] added timm models to test model zoo * polish code * polish code * polish code * polish code * polish code 2023-03-14 06:29:18 +00:00

			`class ModelZooRegistry(dict):`
			`"""`
			`A registry to map model names to model and data generation functions.`
			`"""`

[misc] update pre-commit and run all files (#4752) * [misc] update pre-commit * [misc] run pre-commit * [misc] remove useless configuration files * [misc] ignore cuda for clang-format 2023-09-19 06:20:26 +00:00			`def register(`
			`self,`
			`name: str,`
			`model_fn: Callable,`
			`data_gen_fn: Callable,`
			`output_transform_fn: Callable,`
			`loss_fn: Callable = None,`
			`model_attribute: ModelAttribute = None,`
			`):`
[test] added timm models to test model zoo (#3129) * [test] added timm models to test model zoo * polish code * polish code * polish code * polish code * polish code 2023-03-14 06:29:18 +00:00			`"""`
			`Register a model and data generation function.`

			`Examples:`
[shardformer] adapted T5 and LLaMa test to use kit (#4049) * [shardformer] adapted T5 and LLaMa test to use kit * polish code 2023-06-21 01:32:46 +00:00
			```python
			`# normal forward workflow`
			`model = resnet18()`
			`data = resnet18_data_gen()`
			`output = model(**data)`
			`transformed_output = output_transform_fn(output)`
			`loss = loss_fn(transformed_output)`

			`# Register`
			`model_zoo = ModelZooRegistry()`
			`model_zoo.register('resnet18', resnet18, resnet18_data_gen, output_transform_fn, loss_fn)`
			```
[test] added timm models to test model zoo (#3129) * [test] added timm models to test model zoo * polish code * polish code * polish code * polish code * polish code 2023-03-14 06:29:18 +00:00
			`Args:`
			`name (str): Name of the model.`
[shardformer] adapted T5 and LLaMa test to use kit (#4049) * [shardformer] adapted T5 and LLaMa test to use kit * polish code 2023-06-21 01:32:46 +00:00			`model_fn (Callable): A function that returns a model. It must not contain any arguments.`
			`data_gen_fn (Callable): A function that returns a data sample in the form of Dict. It must not contain any arguments.`
			`output_transform_fn (Callable): A function that transforms the output of the model into Dict.`
			`loss_fn (Callable): a function to compute the loss from the given output. Defaults to None`
[test] added timm models to test model zoo (#3129) * [test] added timm models to test model zoo * polish code * polish code * polish code * polish code * polish code 2023-03-14 06:29:18 +00:00			`model_attribute (ModelAttribute): Attributes of the model. Defaults to None.`
			`"""`
[shardformer] adapted T5 and LLaMa test to use kit (#4049) * [shardformer] adapted T5 and LLaMa test to use kit * polish code 2023-06-21 01:32:46 +00:00			`self[name] = (model_fn, data_gen_fn, output_transform_fn, loss_fn, model_attribute)`
[test] added timm models to test model zoo (#3129) * [test] added timm models to test model zoo * polish code * polish code * polish code * polish code * polish code 2023-03-14 06:29:18 +00:00
[workflow] fixed oom tests (#5275) * [workflow] fixed oom tests * polish * polish * polish 2024-01-16 10:55:13 +00:00			`def get_sub_registry(`
			`self, keyword: Union[str, List[str]], exclude: Union[str, List[str]] = None, allow_empty: bool = False`
			`):`
[test] added timm models to test model zoo (#3129) * [test] added timm models to test model zoo * polish code * polish code * polish code * polish code * polish code 2023-03-14 06:29:18 +00:00			`"""`
			`Get a sub registry with models that contain the keyword.`

			`Args:`
			`keyword (str): Keyword to filter models.`
			`"""`
			`new_dict = dict()`

[workflow] fixed build CI (#5240) * [workflow] fixed build CI * polish * polish * polish * polish * polish 2024-01-10 14:34:16 +00:00			`if isinstance(keyword, str):`
			`keyword_list = [keyword]`
			`else:`
			`keyword_list = keyword`
			`assert isinstance(keyword_list, (list, tuple))`

[ci] fixed ddp test (#5254) * [ci] fixed ddp test * polish 2024-01-11 09:16:32 +00:00			`if exclude is None:`
			`exclude_keywords = []`
			`elif isinstance(exclude, str):`
			`exclude_keywords = [exclude]`
			`else:`
			`exclude_keywords = exclude`
			`assert isinstance(exclude_keywords, (list, tuple))`

[test] added timm models to test model zoo (#3129) * [test] added timm models to test model zoo * polish code * polish code * polish code * polish code * polish code 2023-03-14 06:29:18 +00:00			`for k, v in self.items():`
[workflow] fixed build CI (#5240) * [workflow] fixed build CI * polish * polish * polish * polish * polish 2024-01-10 14:34:16 +00:00			`for kw in keyword_list:`
			`if kw in k:`
[ci] fixed ddp test (#5254) * [ci] fixed ddp test * polish 2024-01-11 09:16:32 +00:00			`should_exclude = False`
			`for ex_kw in exclude_keywords:`
			`if ex_kw in k:`
			`should_exclude = True`

			`if not should_exclude:`
			`new_dict[k] = v`
[shardformer] added embedding gradient check (#4124) 2023-06-30 08:16:44 +00:00
[workflow] fixed oom tests (#5275) * [workflow] fixed oom tests * polish * polish * polish 2024-01-16 10:55:13 +00:00			`if not allow_empty:`
			`assert len(new_dict) > 0, f"No model found with keyword {keyword}"`
[test] added timm models to test model zoo (#3129) * [test] added timm models to test model zoo * polish code * polish code * polish code * polish code * polish code 2023-03-14 06:29:18 +00:00			`return new_dict`


[example]add gpt2 benchmark example script. (#5295) * benchmark gpt2 * fix fix fix fix * [doc] fix typo in Colossal-LLaMA-2/README.md (#5247) * [workflow] fixed build CI (#5240) * [workflow] fixed build CI * polish * polish * polish * polish * polish * [ci] fixed booster test (#5251) * [ci] fixed booster test * [ci] fixed booster test * [ci] fixed booster test * [ci] fixed ddp test (#5254) * [ci] fixed ddp test * polish * fix typo in applications/ColossalEval/README.md (#5250) * [ci] fix shardformer tests. (#5255) * fix ci fix * revert: revert p2p * feat: add enable_metadata_cache option * revert: enable t5 tests --------- Co-authored-by: Wenhao Chen <cwher@outlook.com> * [doc] fix doc typo (#5256) * [doc] fix annotation display * [doc] fix llama2 doc * [hotfix]: add pp sanity check and fix mbs arg (#5268) * fix: fix misleading mbs arg * feat: add pp sanity check * fix: fix 1f1b sanity check * [workflow] fixed incomplete bash command (#5272) * [workflow] fixed oom tests (#5275) * [workflow] fixed oom tests * polish * polish * polish * [ci] fix test_hybrid_parallel_plugin_checkpoint_io.py (#5276) * fix ci fix * fix test * revert: revert p2p * feat: add enable_metadata_cache option * revert: enable t5 tests * fix --------- Co-authored-by: Wenhao Chen <cwher@outlook.com> * [shardformer] hybridparallelplugin support gradients accumulation. (#5246) * support gradients acc fix fix fix fix fix fix fix fix fix fix fix fix fix * fix fix * fix fix fix * [hotfix] Fix ShardFormer test execution path when using sequence parallelism (#5230) * fix auto loading gpt2 tokenizer (#5279) * [doc] add llama2-13B disyplay (#5285) * Update README.md * fix 13b typo --------- Co-authored-by: binmakeswell <binmakeswell@gmail.com> * fix llama pretrain (#5287) * fix * fix * fix fix * fix fix fix * fix fix * benchmark gpt2 * fix fix fix fix * [workflow] fixed build CI (#5240) * [workflow] fixed build CI * polish * polish * polish * polish * polish * [ci] fixed booster test (#5251) * [ci] fixed booster test * [ci] fixed booster test * [ci] fixed booster test * fix fix * fix fix fix * fix * fix fix fix fix fix * fix * Update shardformer.py --------- Co-authored-by: digger yu <digger-yu@outlook.com> Co-authored-by: Frank Lee <somerlee.9@gmail.com> Co-authored-by: Wenhao Chen <cwher@outlook.com> Co-authored-by: binmakeswell <binmakeswell@gmail.com> Co-authored-by: Zhongkai Zhao <kanezz620@gmail.com> Co-authored-by: Michelle <97082656+MichelleMa8@users.noreply.github.com> Co-authored-by: Desperado-Jia <502205863@qq.com> 2024-03-04 08:18:13 +00:00			`model_zoo = ModelZooRegistry()`