[GSoC] DataFrame project bonding
Wei Wang via ublas <[email protected]> Mon, 6 May 2019 21:14:04 -0700
| Newsgroups | gmane.comp.lib.boost.ublas |
|---|---|
| Message-ID | <CAEu7EHEkThVoNZ3CKfyiruTOUCzO1sx7kbVqFax3L8KbMrN2RA@mail.gmail.com> |
--===============7612517596535775210== Content-Type: multipart/alternative; boundary="000000000000300ca505884473f8" --000000000000300ca505884473f8 Content-Type: text/plain; charset="UTF-8" Hi, My name is Wei Wang, and I am lucky to be able to as the student for GSoC19 Boost::ublas library. My project is to build a data structure work like pandas.DataFrame or dataframe in R. As I found on GSoC's website, my mentor will be Bellot, and I'm very glad. I have two questions related to logistics: (1) Where should I work for the code? I find an empty organization in Github(https://github.com/BoostGSoC19), but I'm still not sure how I gonna submit them. (2) Should I fork the whole ublas project? Or simply start build my own project directly under boost/numeric/ublas? Another two questions related to project requirement: (1) I have read one implementation from one previous student ( https://github.com/BoostGSoC17/data_frame), which is pretty good. But it somehow goes against my idea. Is it okay to start a new project? And also I'd like to ask what's your expectation from this project? I'm targeting at pandas.DataFrame(though it won't be that full-featured), but the basics are: - indexing - slicing - sort based on col - relation ops like select, join - set operations on rows like union, set diff, intersect - group (possibly) (2) What should I show in my final submit? Will it be evaluated on whether my code is able to merge? Or simply I will be provided some test case and see if I can pass them? Cheers, Wei --000000000000300ca505884473f8 Content-Type: text/html; charset="UTF-8" Content-Transfer-Encoding: quoted-printable <div dir=3D"ltr">Hi,=C2=A0<div>My name is Wei Wang, and I am lucky to be ab= le to as the student for GSoC19 Boost::ublas library. My project is to buil= d a data structure work like pandas.DataFrame or dataframe in R. As I found= on GSoC's website, my mentor will be Bellot, and I'm very glad.=C2= =A0</div><div>I have two questions related to logistics:=C2=A0</div><div>(1= ) Where should I work for the code? I find an empty organization in Github(= <a href=3D"https://github.com/BoostGSoC19">https://github.com/BoostGSoC19</= a>), but I'm still not sure how I gonna submit them.=C2=A0</div><div>(2= ) Should I fork the whole ublas project? Or simply start build my own proje= ct directly under boost/numeric/ublas?=C2=A0</div><div>Another two question= s related to project requirement:=C2=A0</div><div>(1) I have read one imple= mentation from one previous student (<a href=3D"https://github.com/BoostGSo= C17/data_frame">https://github.com/BoostGSoC17/data_frame</a>), which is pr= etty good. But it somehow goes against my idea. Is it okay to start a new p= roject?=C2=A0</div><div>And also I'd like to ask what's your expect= ation from this project?=C2=A0</div><div>I'm targeting at pandas.DataFr= ame(though it won't be that full-featured), but the basics are:</div><d= iv>- indexing</div><div>- slicing</div><div>- sort based on col</div><div>-= relation ops like select, join</div><div>- set operations on rows like uni= on, set diff, intersect</div><div>- group (possibly)</div><div>(2) What sho= uld I show in my final submit? Will it be evaluated on whether my code is a= ble to merge? Or simply I will be provided some test case and see if I can = pass them?=C2=A0</div><div><br></div><div>Cheers,=C2=A0<br></div><div>Wei</= div></div> --000000000000300ca505884473f8-- --===============7612517596535775210== Content-Type: text/plain; charset="us-ascii" MIME-Version: 1.0 Content-Transfer-Encoding: 7bit Content-Disposition: inline