Repository navigation
Conversation
|
@lmeyerov Despite this being pretty much a work in progress, I already opened a PR to get some initial feedback. I hope that's okay :) This is a partial implementation of table reductions. The function g = graphistry.edges(pd.DataFrame([{'x': 'a', 'y': 'm'},
{'x': 'b', 'y': 'm'},
{'x': 'c', 'y': 'n'},
{'x': 'd', 'y': 'm'}]), 'x', 'y')
g.replace_nodes_with_edges(['m',])
print(g._edges)
#result
x y reduced_from
0 a b m
1 a d m
2 b d m
3 c n NaNWhat are your thoughts? |
|
Cool, some thoughts:
Ex: annoying_nodes_df = g._nodes[ g._nodes['type'] == 'event' ]
g2 = g.replace_nodes_with_edges(annoying_nodes_df)
I'll need to look at |
|
@lmeyerov Thank you very much for your feedback. I have pushed another commit now. Kindly have a look at the For now I would like to make some headway into getting the table reductions into a good shape before starting to work on pregel/graphx style mapping. |
|
Great! In case it helped, I sketched out a few the top use cases we see in the original ticket: #193 re: |
|
Just checking in on this. Maybe the thing to do is:
I'm thinking we can land this as-is, and update the underlying feature request for future PRs to explore (a) a custom reductions mode for this (ex: summary stats) and (b) multi-node collapsing/grouping (ex: collapse connected nodes name/email/address/username into a single labeled node) |
…bench #283) The docs policy caps graphistry/compute commit drift at 12 since a published run's measured commit; this branch adds 8 on top of master's 8. Instead of a waiver, the GraphBench 20k/100k and SNB boards were re-measured at 2d64913 under the published protocol and republished from pyg-bench (#283, same contract v3). Drift for the four re-measured runs is now 0; the bench-provenance directives name the new run ids. Co-Authored-By: Claude Fable 5.1 <[email protected]> Claude-Session: https://claude.ai/code/session_017ropeBMLJUuy6ViYwy15ud
For issue #193