bart-large-nl2sql / README.md

Update README.md

9be8c78 verified over 1 year ago

5.95 kB

	---
	# For reference on model card metadata, see the spec: https://github.com/huggingface/hub-docs/blob/main/modelcard.md?plain=1
	# Doc / guide: https://huggingface.co/docs/hub/model-cards
	{}

	---

	---
	widget:
	- text: "Is this review positive or negative? Review: Best cast iron skillet you will ever buy."
	example_title: "Sentiment analysis"
	- text: "Barack Obama nominated Hilary Clinton as his secretary of state on Monday. He chose her because she had ..."
	example_title: "Coreference resolution"
	- text: "On a shelf, there are five books: a gray book, a red book, a purple book, a blue book, and a black book ..."
	example_title: "Logic puzzles"
	- text: "The two men running to become New York City's next mayor will face off in their first debate Wednesday night ..."
	example_title: "Reading comprehension"
	---

	# Model Card for Model ID

	Generate SQL from Natural Language question with a SQL context.


	## Model Details

	### Model Description

	BART from facebook/bart-large-cnn is fintuned on gretelai/synthetic_text_to_sql dataset to generate SQL from NL and SQL context


	- Model type: [BART]
	- Language(s) (NLP): English
	- License: openrail
	- Finetuned from model [facebook/bart-large-cnn](https://huggingface.co/facebook/bart-large-cnn?text=The+tower+is+324+metres+%281%2C063+ft%29+tall%2C+about+the+same+height+as+an+81-storey+building%2C+and+the+tallest+structure+in+Paris.+Its+base+is+square%2C+measuring+125+metres+%28410+ft%29+on+each+side.+During+its+construction%2C+the+Eiffel+Tower+surpassed+the+Washington+Monument+to+become+the+tallest+man-made+structure+in+the+world%2C+a+title+it+held+for+41+years+until+the+Chrysler+Building+in+New+York+City+was+finished+in+1930.+It+was+the+first+structure+to+reach+a+height+of+300+metres.+Due+to+the+addition+of+a+broadcasting+aerial+at+the+top+of+the+tower+in+1957%2C+it+is+now+taller+than+the+Chrysler+Building+by+5.2+metres+%2817+ft%29.+Excluding+transmitters%2C+the+Eiffel+Tower+is+the+second+tallest+free-standing+structure+in+France+after+the+Millau+Viaduct.)
	- Dataset: [gretelai/synthetic_text_to_sql](https://huggingface.co/datasets/gretelai/synthetic_text_to_sql)

	## Uses

	<!-- Address questions around how the model is intended to be used, including the foreseeable users of the model and those affected by the model. -->

	### Direct Use

	<!-- This section is for the model use without fine-tuning or plugging into a larger ecosystem/app. -->

	[More Information Needed]

	### Downstream Use [optional]

	<!-- This section is for the model use when fine-tuned for a task, or when plugged into a larger ecosystem/app -->

	[More Information Needed]


	## How to Get Started with the Model

	Use the code below to get started with the model.

	[More Information Needed]

	## Training Details

	### Training Data

	<!-- This should link to a Dataset Card, perhaps with a short stub of information on what the training data is all about as well as documentation related to data pre-processing or additional filtering. -->

	[More Information Needed]

	### Training Procedure

	<!-- This relates heavily to the Technical Specifications. Content here should link to that section when it is relevant to the training procedure. -->

	#### Preprocessing [optional]

	[More Information Needed]


	#### Training Hyperparameters

	- Training regime: [More Information Needed] <!--fp32, fp16 mixed precision, bf16 mixed precision, bf16 non-mixed precision, fp16 non-mixed precision, fp8 mixed precision -->

	#### Speeds, Sizes, Times [optional]

	<!-- This section provides information about throughput, start/end time, checkpoint size if relevant, etc. -->

	[More Information Needed]

	## Evaluation

	<!-- This section describes the evaluation protocols and provides the results. -->

	### Testing Data, Factors & Metrics

	#### Testing Data

	<!-- This should link to a Dataset Card if possible. -->

	[More Information Needed]

	#### Factors

	<!-- These are the things the evaluation is disaggregating by, e.g., subpopulations or domains. -->

	[More Information Needed]

	#### Metrics

	<!-- These are the evaluation metrics being used, ideally with a description of why. -->

	[More Information Needed]

	### Results

	[More Information Needed]


	## Technical Specifications [optional]

	### Model Architecture and Objective

	[More Information Needed]

	### Compute Infrastructure

	[More Information Needed]

	#### Hardware

	[More Information Needed]


	## Citation

	<!-- If there is a paper or blog post introducing the model, the APA and Bibtex information for that should go in this section. -->

	BibTeX:

	@software{gretel-synthetic-text-to-sql-2024,
	author = {Meyer, Yev and Emadi, Marjan and Nathawani, Dhruv and Ramaswamy, Lipika and Boyd, Kendrick and Van Segbroeck, Maarten and Grossman, Matthew and Mlocek, Piotr and Newberry, Drew},
	title = {{Synthetic-Text-To-SQL}: A synthetic dataset for training language models to generate SQL queries from natural language prompts},
	month = {April},
	year = {2024},
	url = {https://huggingface.co/datasets/gretelai/synthetic-text-to-sql}
	}


	@article{DBLP:journals/corr/abs-1910-13461,
	author = {Mike Lewis and
	Yinhan Liu and
	Naman Goyal and
	Marjan Ghazvininejad and
	Abdelrahman Mohamed and
	Omer Levy and
	Veselin Stoyanov and
	Luke Zettlemoyer},
	title = {{BART:} Denoising Sequence-to-Sequence Pre-training for Natural Language
	Generation, Translation, and Comprehension},
	journal = {CoRR},
	volume = {abs/1910.13461},
	year = {2019},
	url = {http://arxiv.org/abs/1910.13461},
	eprinttype = {arXiv},
	eprint = {1910.13461},
	timestamp = {Thu, 31 Oct 2019 14:02:26 +0100},
	biburl = {https://dblp.org/rec/journals/corr/abs-1910-13461.bib},
	bibsource = {dblp computer science bibliography, https://dblp.org}
	}


	## Model Card Authors [optional]

	[Swastik Maiti]