Name		Name	Last commit message	Last commit date
Latest commit History 41 Commits
.github		.github
src/tasknet		src/tasknet
CITATION.cff		CITATION.cff
LICENSE		LICENSE
README.md		README.md
Tasknet_multitask_transformer_training.ipynb		Tasknet_multitask_transformer_training.ipynb
pyproject.toml		pyproject.toml
setup.cfg		setup.cfg

Repository files navigation

tasknet

tasknet is an interface between Huggingface datasets and Huggingface Trainer.

Task templates

tasknet relies on task templates to avoid boilerplate codes. The task templates are correspond to Transformers AutoClasses:

SequenceClassification
TokenClassification
MultipleChoice

The task templates follow the same interface. They implement preprocess_function and compute_metrics. Look at tasks.py and use existing templates as a starting point to implement a custom task template.

Instanciating a task

Each task template is associated with specific fields. Classification has two text fields s1,s2, and a label y. Pass a dataset to a template, and fill-in the mapping between the dataset fields and the template fields to instanciate a task.

import tasknet as tn
from datasets import load_dataset

rte = tn.Classification(
    dataset=load_dataset("glue", "rte"),
    s1="sentence1", s2="sentence2", y="label"
)

class args:
  model_name='roberta-base'
  learning_rate = 3e-5 # see https://huggingface.co/docs/transformers/v4.24.0/en/main_classes/trainer#transformers.TrainingArguments

 
tasks = [rte]
model = tn.Model(tasks, args)
trainer = tn.Trainer(model, tasks, args)
trainer.train()

As you can see, tasknet is multitask by design. It works with list of tasks and the model creates a task_models_list attribute.

Installation

pip install tasknet

Additional examples:

Colab:

https://colab.research.google.com/drive/15Xf4Bgs3itUmok7XlAK6EEquNbvjD9BD?usp=sharing

tasknet vs jiant

jiant is another library comparable to tasknet. tasknet is a minimal extension of Trainer centered on task templates, while jiant builds a custom analog of Trainer from scratch called runner. tasknet is leaner and easier to extend. jiant is config-based while tasknet is designed for interative use and scripting.

Credit

This code uses some part of the examples of the transformers library and some code from multitask-learning-transformers.

Contact

You can request features on github or reach me at [email protected]

@misc{sileod21-tasknet,
  author = {Sileo, Damien},
  doi = {10.5281/zenodo.561225781},
  month = {11},
  title = {{tasknet, multitask interface between Trainer and datasets}},
  url = {https://github.com/sileod/tasknet},
  version = {1.5.0},
  year = {2022}}

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

Repository files navigation

tasknet

Task templates

Instanciating a task

Installation

Additional examples:

Colab:

tasknet vs jiant

Credit

Contact

About

Releases 59

Packages

Contributors 2

Languages

License

sileod/tasknet

Folders and files

Latest commit

History

Repository files navigation

tasknet

Task templates

Instanciating a task

Installation

Additional examples:

Colab:

tasknet vs jiant

Credit

Contact

About

Topics

Resources

License

Citation

Stars

Watchers

Forks

Releases 59

Packages 0

Contributors 2

Languages

Packages