Thank you for publishing the model. Recently, after reading your article, I saw in the code that you downloaded the dataset "1 Bill Word Language Model Benchmark". This dataset seems to be unrelated to the Escherichia coli promoter. What is the purpose of this dataset?
I also want to ask you a question. Do you directly process the startup Subsequence into one hot vector in the code and generate it with GAN? Is there a dimensionality increase operation on the one-hot matrix?
Looking forward to your reply!
Thank you for publishing the model. Recently, after reading your article, I saw in the code that you downloaded the dataset "1 Bill Word Language Model Benchmark". This dataset seems to be unrelated to the Escherichia coli promoter. What is the purpose of this dataset?
I also want to ask you a question. Do you directly process the startup Subsequence into one hot vector in the code and generate it with GAN? Is there a dimensionality increase operation on the one-hot matrix?
Looking forward to your reply!