SwastikM commited on
Commit
e624496
1 Parent(s): fee4094

Create README.md

Browse files
Files changed (1) hide show
  1. README.md +159 -0
README.md ADDED
@@ -0,0 +1,159 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ ---
2
+ # For reference on model card metadata, see the spec: https://github.com/huggingface/hub-docs/blob/main/modelcard.md?plain=1
3
+ # Doc / guide: https://huggingface.co/docs/hub/model-cards
4
+ {}
5
+ ---
6
+
7
+ # Model Card for Model ID
8
+
9
+ Generate SQL from Natural Language question with a SQL context.
10
+
11
+
12
+ ## Model Details
13
+
14
+ ### Model Description
15
+
16
+ BART from facebook/bart-large-cnn is fintuned on gretelai/synthetic_text_to_sql dataset to generate SQL from NL and SQL context
17
+
18
+
19
+ - **Model type:** [BART]
20
+ - **Language(s) (NLP):** English
21
+ - **License:** openrail
22
+ - **Finetuned from model [facebook/bart-large-cnn](https://huggingface.co/facebook/bart-large-cnn?text=The+tower+is+324+metres+%281%2C063+ft%29+tall%2C+about+the+same+height+as+an+81-storey+building%2C+and+the+tallest+structure+in+Paris.+Its+base+is+square%2C+measuring+125+metres+%28410+ft%29+on+each+side.+During+its+construction%2C+the+Eiffel+Tower+surpassed+the+Washington+Monument+to+become+the+tallest+man-made+structure+in+the+world%2C+a+title+it+held+for+41+years+until+the+Chrysler+Building+in+New+York+City+was+finished+in+1930.+It+was+the+first+structure+to+reach+a+height+of+300+metres.+Due+to+the+addition+of+a+broadcasting+aerial+at+the+top+of+the+tower+in+1957%2C+it+is+now+taller+than+the+Chrysler+Building+by+5.2+metres+%2817+ft%29.+Excluding+transmitters%2C+the+Eiffel+Tower+is+the+second+tallest+free-standing+structure+in+France+after+the+Millau+Viaduct.)**
23
+ - **Dataset:** [gretelai/synthetic_text_to_sql](https://huggingface.co/datasets/gretelai/synthetic_text_to_sql)
24
+
25
+ ## Uses
26
+
27
+ <!-- Address questions around how the model is intended to be used, including the foreseeable users of the model and those affected by the model. -->
28
+
29
+ ### Direct Use
30
+
31
+ <!-- This section is for the model use without fine-tuning or plugging into a larger ecosystem/app. -->
32
+
33
+ [More Information Needed]
34
+
35
+ ### Downstream Use [optional]
36
+
37
+ <!-- This section is for the model use when fine-tuned for a task, or when plugged into a larger ecosystem/app -->
38
+
39
+ [More Information Needed]
40
+
41
+
42
+ ## How to Get Started with the Model
43
+
44
+ Use the code below to get started with the model.
45
+
46
+ [More Information Needed]
47
+
48
+ ## Training Details
49
+
50
+ ### Training Data
51
+
52
+ <!-- This should link to a Dataset Card, perhaps with a short stub of information on what the training data is all about as well as documentation related to data pre-processing or additional filtering. -->
53
+
54
+ [More Information Needed]
55
+
56
+ ### Training Procedure
57
+
58
+ <!-- This relates heavily to the Technical Specifications. Content here should link to that section when it is relevant to the training procedure. -->
59
+
60
+ #### Preprocessing [optional]
61
+
62
+ [More Information Needed]
63
+
64
+
65
+ #### Training Hyperparameters
66
+
67
+ - **Training regime:** [More Information Needed] <!--fp32, fp16 mixed precision, bf16 mixed precision, bf16 non-mixed precision, fp16 non-mixed precision, fp8 mixed precision -->
68
+
69
+ #### Speeds, Sizes, Times [optional]
70
+
71
+ <!-- This section provides information about throughput, start/end time, checkpoint size if relevant, etc. -->
72
+
73
+ [More Information Needed]
74
+
75
+ ## Evaluation
76
+
77
+ <!-- This section describes the evaluation protocols and provides the results. -->
78
+
79
+ ### Testing Data, Factors & Metrics
80
+
81
+ #### Testing Data
82
+
83
+ <!-- This should link to a Dataset Card if possible. -->
84
+
85
+ [More Information Needed]
86
+
87
+ #### Factors
88
+
89
+ <!-- These are the things the evaluation is disaggregating by, e.g., subpopulations or domains. -->
90
+
91
+ [More Information Needed]
92
+
93
+ #### Metrics
94
+
95
+ <!-- These are the evaluation metrics being used, ideally with a description of why. -->
96
+
97
+ [More Information Needed]
98
+
99
+ ### Results
100
+
101
+ [More Information Needed]
102
+
103
+
104
+ ## Technical Specifications [optional]
105
+
106
+ ### Model Architecture and Objective
107
+
108
+ [More Information Needed]
109
+
110
+ ### Compute Infrastructure
111
+
112
+ [More Information Needed]
113
+
114
+ #### Hardware
115
+
116
+ [More Information Needed]
117
+
118
+
119
+ ## Citation
120
+
121
+ <!-- If there is a paper or blog post introducing the model, the APA and Bibtex information for that should go in this section. -->
122
+
123
+ **BibTeX:**
124
+
125
+ @software{gretel-synthetic-text-to-sql-2024,
126
+ author = {Meyer, Yev and Emadi, Marjan and Nathawani, Dhruv and Ramaswamy, Lipika and Boyd, Kendrick and Van Segbroeck, Maarten and Grossman, Matthew and Mlocek, Piotr and Newberry, Drew},
127
+ title = {{Synthetic-Text-To-SQL}: A synthetic dataset for training language models to generate SQL queries from natural language prompts},
128
+ month = {April},
129
+ year = {2024},
130
+ url = {https://huggingface.co/datasets/gretelai/synthetic-text-to-sql}
131
+ }
132
+
133
+
134
+ @article{DBLP:journals/corr/abs-1910-13461,
135
+ author = {Mike Lewis and
136
+ Yinhan Liu and
137
+ Naman Goyal and
138
+ Marjan Ghazvininejad and
139
+ Abdelrahman Mohamed and
140
+ Omer Levy and
141
+ Veselin Stoyanov and
142
+ Luke Zettlemoyer},
143
+ title = {{BART:} Denoising Sequence-to-Sequence Pre-training for Natural Language
144
+ Generation, Translation, and Comprehension},
145
+ journal = {CoRR},
146
+ volume = {abs/1910.13461},
147
+ year = {2019},
148
+ url = {http://arxiv.org/abs/1910.13461},
149
+ eprinttype = {arXiv},
150
+ eprint = {1910.13461},
151
+ timestamp = {Thu, 31 Oct 2019 14:02:26 +0100},
152
+ biburl = {https://dblp.org/rec/journals/corr/abs-1910-13461.bib},
153
+ bibsource = {dblp computer science bibliography, https://dblp.org}
154
+ }
155
+
156
+
157
+ ## Model Card Authors [optional]
158
+
159
+ [Swastik Maiti]